AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,059 results
21 Apr 2026

E2E-GMNER: End-to-End Generative Grounded Multimodal Named Entity Recognition

ResearchDGX agent

arXiv:2604.17319v1 Announce Type: cross Abstract: Grounded Multimodal Named Entity Recognition (GMNER) aims to jointly identify named entity mentions in text, predict their semantic types, and ground

Efficient Task Adaptation in Large Language Models via Selective Parameter Optimization

Model ReleasesDGX agent

arXiv:2604.17051v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated excellent performance in general language understanding, generation and other tasks. However, when fine-t

Enabling Stroke-Level Structural Analysis of Hieroglyphic Scripts without Language-Specific Priors

ResearchDGX agent

arXiv:2601.05508v2 Announce Type: replace-cross Abstract: Hieroglyphs, as logographic writing systems, encode rich semantic and cultural information within their internal structural composition. Yet,

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Enhancing Anomaly-Based Intrusion Detection Systems with Process Mining

ResearchDGX agent

arXiv:2604.18066v1 Announce Type: cross Abstract: Anomaly-based Intrusion Detection Systems (IDSs) ensure protection against malicious attacks on networked systems. While deep learning-based IDSs achi

ESsEN: Training Compact Discriminative Vision-Language Transformers in a Low-Resource Setting

Model ReleasesDGX agent

arXiv:2604.18452v1 Announce Type: cross Abstract: Vision-language modeling is rapidly increasing in popularity with an ever expanding list of available models. In most cases, these vision-language mod

Estimating Commonsense Plausibility through Semantic Shifts

ResearchDGX agent

arXiv:2502.13464v2 Announce Type: replace Abstract: Commonsense plausibility estimation is critical for evaluating language models (LMs), yet existing generative approaches--reliant on likelihoods or

Evolutionary Negative Module Pruning for Better LoRA Merging

SafetyDGX agent

arXiv:2604.17753v1 Announce Type: cross Abstract: Merging multiple Low-Rank Adaptation (LoRA) experts into a single backbone is a promising approach for efficient multi-task deployment. While existing

Exploring Mutual Cross-Modal Attention for Context-Aware Human Affordance Generation

ResearchDGX agent

arXiv:2502.13637v2 Announce Type: replace Abstract: Human affordance learning investigates contextually relevant novel pose prediction such that the estimated pose represents a valid human action with

Faithfulness vs. Safety: Evaluating LLM Behavior Under Counterfactual Medical Evidence

SafetyDGX agent

arXiv:2601.11886v2 Announce Type: replace Abstract: In high-stakes domains like medicine, it may be generally desirable for models to faithfully adhere to the context provided. But what happens if the

FaithLens: Detecting and Explaining Faithfulness Hallucination

Model ReleasesDGX agent

arXiv:2512.20182v3 Announce Type: replace Abstract: Recognizing whether outputs from large language models (LLMs) contain faithfulness hallucination is crucial for real-world applications, e.g., retri

Federation over Text: Insight Sharing for Multi-Agent Reasoning

AgentsDGX agent

arXiv:2604.16778v1 Announce Type: new Abstract: LLM-powered agents often reason from scratch when presented with a new problem instance and lack automatic mechanisms to transfer learned skills to othe

FedOBP: Federated Optimal Brain Personalization through Cloud-Edge Element-wise Decoupling

Model ReleasesDGX agent

arXiv:2604.16574v1 Announce Type: new Abstract: Federated Learning (FL) faces challenges from client data heterogeneity and resource-constrained mobile devices, which can degrade model accuracy. Perso

FLARE: A Data-Efficient Surrogate for Predicting Displacement Fields in Directed Energy Deposition

Model ReleasesDGX agent

arXiv:2604.16649v1 Announce Type: new Abstract: Directed energy deposition (DED) produces complex thermo-mechanical responses that can lead to distortion and reduced dimensional accuracy of a manufact

Flow-Opt: Scalable Centralized Multi-Robot Trajectory Optimization with Flow Matching and Differentiable Optimization

SafetyDGX agent

arXiv:2510.09204v2 Announce Type: replace-cross Abstract: Centralized trajectory optimization in the joint space of multiple robots allows access to a larger feasible space that can result in smoother

FlowRefiner: Flow Matching-Based Iterative Refinement for 3D Turbulent Flow Simulation

ResearchDGX agent

arXiv:2604.17149v1 Announce Type: cross Abstract: Accurate autoregressive prediction of 3D turbulent flows remains challenging for neural PDE solvers, as small errors in fine-scale structures can accu

FM-CAC: Carbon-Aware Control for Battery-Buffered Edge AI via Time-Series Foundation Models

Local AiDGX agent

arXiv:2604.16448v1 Announce Type: cross Abstract: As edge AI deployments scale to billions of devices running always-on, real-time compound AI pipelines, they represent a massive and largely unmanaged

FoleyDirector: Fine-Grained Temporal Steering for Video-to-Audio Generation via Structured Scripts

ResearchDGX agent

arXiv:2603.19857v2 Announce Type: replace-cross Abstract: Recent Video-to-Audio (V2A) methods have achieved remarkable progress, enabling the synthesis of realistic, high-quality audio. However, they

Forest Before Trees: Latent Superposition for Efficient Visual Reasoning

Local AiDGX agent

arXiv:2601.06803v2 Announce Type: replace Abstract: While Chain-of-Thought empowers Large Vision-Language Models with multi-step reasoning, explicit textual rationales suffer from an information bandw

Fractal Characterization of Low-Correlation Signals in AI-Generated Image Detection

ResearchDGX agent

arXiv:2604.17268v1 Announce Type: new Abstract: AI-generated imagery has reached near-photorealistic fidelity, yet this technology poses significant threats to information security and societal trust.

Frequency-guided Multi-level Reasoning for Scene Graph Generation in Video

ResearchDGX agent

arXiv:2604.17298v1 Announce Type: new Abstract: Video Scene Graph Generation aims to obtain structured semantic representations of objects and their relationships in videos for high-level understandin

FSEVAL: Feature Selection Evaluation Toolbox and Dashboard

ResearchDGX agent

arXiv:2604.18227v1 Announce Type: new Abstract: Feature selection is a fundamental machine learning and data mining task, involved with discriminating redundant features from informative ones. It is a

Functional Similarity Metric for Neural Networks: Overcoming Parametric Ambiguity via Activation Region Analysis

Model ReleasesDGX agent

arXiv:2604.16426v1 Announce Type: new Abstract: As modern deep learning architectures grow in complexity, representational ambiguity emerges as a critical barrier to their interpretability and reliabl

Generative Semantic Communication via Alternating Dual-Domain Posterior Sampling

SafetyDGX agent

arXiv:2604.16796v1 Announce Type: new Abstract: Generative semantic communication (SemCom) harnesses pretrained generative priors to improve the perceptual quality of wireless image transmission. Exis

Geodesic Semantic Search: Cartographic Navigation of Citation Graphs with Learned Local Riemannian Maps

Local AiDGX agent

arXiv:2602.23665v3 Announce Type: replace-cross Abstract: We present Geodesic Semantic Search (GSS), a retrieval system that learns node-specific Riemannian metrics on citation graphs to enable geomet

Geometry-Guided 3D Visual Token Pruning for Video-Language Models

ResearchDGX agent

arXiv:2604.18260v1 Announce Type: new Abstract: Multimodal large language models have demonstrated remarkable capabilities in 2D vision, motivating their extension to 3D scene understanding. Recent st

GeoRC: A Benchmark for Geolocation Reasoning Chains

Model ReleasesDGX agent

arXiv:2601.21278v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are good at recognizing the global location of a photograph -- their geolocation prediction accuracy rivals the

Graph Neural Networks for Graphs with Heterophily: A Survey

ApplicationsDGX agent

arXiv:2202.07082v4 Announce Type: replace Abstract: Recent years have witnessed fast developments of graph neural networks (GNNs) that have benefited myriad graph analytic tasks and applications. Most

GSQ: Highly-Accurate Low-Precision Scalar Quantization for LLMs via Gumbel-Softmax Sampling

Model ReleasesDGX agent

arXiv:2604.18556v1 Announce Type: new Abstract: Weight quantization has become a standard tool for efficient LLM deployment, especially for local inference, where models are now routinely served at 2-

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)!

HardwareDGX agent

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)! Introducing ml-intern, the agent that just automated the post-training team @hugg

High-Resolution Visual Reasoning via Multi-Turn Grounding-Based Reinforcement Learning

SafetyDGX agent

arXiv:2507.05920v2 Announce Type: replace Abstract: State-of-the-art large multi-modal models (LMMs) face challenges when processing high-resolution images, as these inputs are converted into enormous

HopWeaver: Cross-Document Synthesis of High-Quality and Authentic Multi-Hop Questions

ResearchDGX agent

arXiv:2505.15087v3 Announce Type: replace Abstract: Multi-Hop Question Answering (MHQA) is crucial for evaluating the model's capability to integrate information from diverse sources. However, creatin

Horizon-Aware Forecasting of Passenger Assistance Demand for Rail Station Workforce Planning

ApplicationsDGX agent

arXiv:2604.16464v1 Announce Type: cross Abstract: Passenger assistance services are essential for accessible rail travel, yet demand varies substantially across stations and over time, creating challe

How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study

Model ReleasesDGX agent

arXiv:2505.15404v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) have achieved remarkable success on reasoning-intensive tasks such as mathematics and programming. However, their enha

Hybrid Quantum Neural Networks for Enhanced Breast Cancer Thermographic Classification: A Novel Quantum-Classical Integration Approach

ApplicationsDGX agent

arXiv:2604.16953v1 Announce Type: cross Abstract: Breast cancer diagnosis through thermographic image analysis remains a critical challenge in medical AI, with classical deep learning approaches facin

I have been using GPT ImageGen-2 for the past weeks I didn't think that better image-generators would be a big deal but it turns out that th…

ApplicationsDGX agent

I have been using GPT ImageGen-2 for the past weeks I didn't think that better image-generators would be a big deal but it turns out that there is a quality threshold I didn't expect, where you can no

IDOBE: Infectious Disease Outbreak forecasting Benchmark Ecosystem

Model ReleasesDGX agent

arXiv:2604.18521v1 Announce Type: new Abstract: Epidemic forecasting has become an integral part of real-time infectious disease outbreak response. While collaborative ensembles composed of statistica

Improving Speech Recognition of Named Entities in Classroom Speech with LLM Revision and Phonetic-Semantic Context

ResearchDGX agent

arXiv:2506.10779v2 Announce Type: replace Abstract: Classroom speech and lectures often contain named entities (NEs) such as names of people and special terminology. While automatic speech recognition

In Search of Lost DNA Sequence Pretraining

ResearchDGX agent

arXiv:2604.16570v1 Announce Type: new Abstract: DNA sequence encoding is fundamental to gene function prediction, protein synthesis, and diverse downstream biological tasks. Despite the substantial pr

Inference-Time Temporal Probability Smoothing for Stable Video Segmentation with SAM2 under Weak Prompts

ResearchDGX agent

arXiv:2604.17115v1 Announce Type: new Abstract: Interactive video segmentation models such as SAM2 have demonstrated strong generalization across diverse visual domains. However, under weak user super

Infrastructure-Centric World Models: Bridging Temporal Depth and Spatial Breadth for Roadside Perception

SafetyDGX agent

arXiv:2604.17651v1 Announce Type: new Abstract: World models, generative AI systems that simulate how environments evolve, are transforming autonomous driving, yet all existing approaches adopt an ego

Instinct vs. Reflection: Unifying Token and Verbalized Confidence in Multimodal Large Models

SafetyDGX agent

arXiv:2604.17274v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated exceptional capabilities in various perception and reasoning tasks. Despite this success, ens

Integrated Wheel Sensor Communication using ESP32 -- A Contribution towards a Digital Twin of the Road System

SafetyDGX agent

arXiv:2509.04061v2 Announce Type: replace Abstract: While current onboard state estimation methods are adequate for most driving and safety-related applications, they do not provide insights into the

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training …

HardwareDGX agent

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training setup iteration. - Data processing workflows. Huge shoutout

Judge a Book by its Cover: Investigating Multi-Modal LLMs for Multi-Page Handwritten Document Transcription

Model ReleasesDGX agent

arXiv:2502.20295v2 Announce Type: replace-cross Abstract: Handwriting text recognition (HTR) remains a challenging task. Existing approaches require fine-tuning on labeled data, which is impractical t

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is…

HardwareDGX agent

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is just scratching the surface. Ultimately, it boils down to g

LBFTI: Layer-Based Facial Template Inversion for Identity-Preserving Fine-Grained Face Reconstruction

ResearchDGX agent

arXiv:2604.18358v1 Announce Type: new Abstract: In face recognition systems, facial templates are widely adopted for identity authentication due to their compliance with the data minimization principl

Let's talk parsing charts 📊📈. Last week we released ParseBench, the first document OCR benchmark for AI agents. New in ParseBench: ChartDa…

Model ReleasesDGX agent

Let's talk parsing charts 📊📈. Last week we released ParseBench, the first document OCR benchmark for AI agents. New in ParseBench: ChartDataPointMatch. Most document look at a chart and OCR the captio

Lightweight Cybersickness Detection based on User-Specific Eye and Head Tracking Data in Virtual Reality

ApplicationsDGX agent

arXiv:2604.17158v1 Announce Type: cross Abstract: The occurrence of cybersickness in virtual reality (VR) significantly impairs users' perception and sense of immersion. Therefore, timely detection of

Lil: Less is Less When Applying Post-Training Sparse-Attention Algorithms in Long-Decode Stage

ResearchDGX agent

arXiv:2601.03043v3 Announce Type: replace Abstract: Large language models (LLMs) demonstrate strong capabilities across a wide range of complex tasks and are increasingly deployed at scale, placing si

LiquidTAD: An Efficient Method for Temporal Action Detection via Liquid Neural Dynamics

Model ReleasesDGX agent

arXiv:2604.18274v1 Announce Type: new Abstract: Temporal Action Detection (TAD) in untrimmed videos is currently dominated by Transformer-based architectures. While high-performing, their quadratic co

Live LTL Progress Tracking: Towards Task-Based Exploration

AgentsDGX agent

arXiv:2604.17106v1 Announce Type: new Abstract: Motivated by the challenge presented by non-Markovian objectives in reinforcement learning (RL), we present a novel framework to track and represent the

LLaMA-XR: A Novel Framework for Radiology Report Generation using LLaMA and QLoRA Fine Tuning

Model ReleasesDGX agent

arXiv:2506.03178v2 Announce Type: replace-cross Abstract: Automated radiology report generation holds significant potential to reduce radiologists' workload and enhance diagnostic accuracy. However, g

LLM-AUG: Robust Wireless Data Augmentation with In-Context Learning in Large Language Models

ResearchDGX agent

arXiv:2604.17770v1 Announce Type: new Abstract: Data scarcity remains a fundamental bottleneck in applying deep learning to wireless communication problems, particularly in scenarios where collecting

Long-CODE: Isolating Pure Long-Context as an Orthogonal Dimension in Video Evaluation

Model ReleasesDGX agent

arXiv:2604.17428v1 Announce Type: new Abstract: As video generation models achieve unprecedented capabilities, the demand for robust video evaluation metrics becomes increasingly critical. Traditional

Low-rank Orthogonalization for Large-scale Matrix Optimization with Applications to Foundation Model Training

Model ReleasesDGX agent

arXiv:2509.11983v2 Announce Type: replace Abstract: Neural network (NN) training is inherently a large-scale matrix optimization problem, yet the matrix structure of NN parameters has long been overlo

Lower Bounds and Proximally Anchored SGD for Non-Convex Minimization Under Unbounded Variance

ResearchDGX agent

arXiv:2604.16620v1 Announce Type: new Abstract: Analysis of Stochastic Gradient Descent (SGD) and its variants typically relies on the assumption of uniformly bounded variance, a condition that freque

LTRR: Learning To Rank Retrievers for LLMs

ResearchDGX agent

arXiv:2506.13743v2 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) systems typically rely on a single fixed retriever, despite growing evidence that no single retriever performs

MambaKick: Early Penalty Direction Prediction from HAR Embeddings

ApplicationsDGX agent

arXiv:2604.16588v1 Announce Type: new Abstract: Penalty kicks in soccer are decided under extreme time constraints, where goalkeepers benefit from anticipating shot direction from the kickers motion b

MASPO: Unifying Gradient Utilization, Probability Mass, and Signal Reliability for Robust and Sample-Efficient LLM Reasoning

SafetyDGX agent

arXiv:2602.17550v3 Announce Type: replace Abstract: Existing Reinforcement Learning with Verifiable Rewards (RLVR) algorithms, such as GRPO, rely on rigid, uniform, and symmetric trust region mechanis

MEDN: Motion-Emotion Feature Decoupling Network for Micro-Expression Recognition

Model ReleasesDGX agent

arXiv:2604.17899v1 Announce Type: new Abstract: Unlike macro-expression, micro-expression does not follow a strictly consistent mapping rule between emotions and Action Units (AUs). As a result, some

← Previous
1…182183184185186…201
Next →