AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
Human
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
24 Jun 2026

Universal Guideline-Driven Image Clustering via a Hybrid LLM Agent

AgentsDGX agent

arXiv:2606.24094v1 Announce Type: new Abstract: Unifying image clustering across different clustering scenarios remains challenging due to fundamental gaps among tasks. We introduce a Guideline-Driven

UOL@IDEM at BEA 2026 Shared Task 1: Neural Fusion and Feature-Rich Modeling for L1-Aware Vocabulary Difficulty Prediction

SafetyDGX agent

arXiv:2606.24501v1 Announce Type: new Abstract: This paper describes UOL@IDEM's closed-track submission to the BEA 2026 shared task on L1-aware vocabulary difficulty prediction. We model the task as r

Variational Model Merging for Pareto Front Estimation in Multitask Finetuning

Model ReleasesDGX agent

arXiv:2412.08147v2 Announce Type: replace-cross Abstract: Pareto fronts are useful to find good task-mixing strategies for multitask finetuning, but they are also costly to compute. To reduce costs, r

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Varying Bundle Size Reactive Multi-Task Assignment using Selective Cost Estimation for Multi-Agent Systems

AgentsDGX agent

arXiv:2606.24462v1 Announce Type: new Abstract: This paper presents a scalable framework for multi-robot task allocation in complex environments where estimating task execution costs is computationall

Verifiable Foundation Models for Robot Safety

SafetyDGX agent

arXiv:2606.23754v1 Announce Type: cross Abstract: Deploying foundation models for robot control raises a central challenge: the expressive power that enables rich, multimodal perception also makes the

VeriPilot: An LLM-Powered Verilog Debugging Framework

Model ReleasesDGX agent

arXiv:2606.23759v1 Announce Type: cross Abstract: Verilog debugging remains one of the most time-consuming stages in digital circuit design. Recent advances in Large Language Models (LLMs) have enable

VeryTrace: Verifying Reasoning Traces through Compilable Formalism and Structured Verification

ResearchDGX agent

arXiv:2606.24124v1 Announce Type: new Abstract: Multi-step reasoning with Chain-of-Thought (CoT) prompting remains fragile: logical errors or hallucinations in early steps silently propagate, producin

video-SALMONN-R^3: Learning to ReWatch, ReAsk, and ReAnswer for Efficient Video Understanding

Model ReleasesDGX agent

arXiv:2606.24477v1 Announce Type: cross Abstract: Video large language models (LLMs) are often constrained by computation and memory budgets, leading them to use reduced frame rates and spatial resolu

VieSpeaker: A Large-Scale Vietnamese Speaker Recognition Dataset Beyond Visual Dependency

ResearchDGX agent

arXiv:2606.24066v1 Announce Type: cross Abstract: Speaker recognition has advanced rapidly with large-scale training datasets, yet Vietnamese remains under-resourced, with existing corpora limited in

VisChronos: Revolutionizing Image Captioning Through Real-Life Events

ApplicationsDGX agent

arXiv:2606.24058v1 Announce Type: new Abstract: This paper aims to bridge the semantic gap between visual content and natural language understanding by leveraging historical events in the real world a

VisCritic: Visual State Comparison as Process Reward for GUI Agents

Model ReleasesDGX agent

arXiv:2606.24525v1 Announce Type: new Abstract: GUI agents powered by vision-language models show strong potential for automating digital tasks, yet frequently fail in long-horizon scenarios due to th

Vision-Language Model Reasoning for Contextual Semantic Mapping in Intralogistics

AgentsDGX agent

arXiv:2606.24814v1 Announce Type: new Abstract: Autonomous mobile robots operating in intralogistics environments rely on geometric maps for localization and navigation, but lack semantic understandin

VistaRef: Boosting Visual Spatial Orientation Awareness for Pointing-to-Object Detection

Local AiDGX agent

arXiv:2606.24498v1 Announce Type: new Abstract: Grounding deictic gestures in natural images is fundamental to AR and human-robot collaboration, providing a basis for seamless spatial interaction. Whi

Visualizing 'We the People': Bridging the Perception Gap through Pluralistic Data Storytelling

ResearchDGX agent

arXiv:2606.24635v1 Announce Type: cross Abstract: Traditional visual data storytelling relies on binary graphics that depict two simplified groups in conflict. This can increase political polarization

ViTexQA: A Multi-Frame Temporal Perception Dataset for Video Text Question Answering

ApplicationsDGX agent

arXiv:2606.24602v1 Announce Type: new Abstract: Despite remarkable progress in multimodal understanding, current MLLMs still exhibit limitations in video text understanding, particularly when semantic

VoltanaLLM: Energy-Efficient and SLO-Aware Disaggregated LLM Serving via Adaptive Frequency Control and State-Space Routing

HardwareDGX agent

arXiv:2509.04827v3 Announce Type: replace-cross Abstract: The energy cost of Large Language Model (LLM) inference is rapidly becoming a barrier to sustainable and scalable deployment. Although modern

VSANet: View-aware Sparse Attention Network for Light Field Image Denoising

ResearchDGX agent

arXiv:2606.24737v1 Announce Type: new Abstract: Light field (LF) image denoising is challenging due to the high-dimensional structure of LF data. While noise is independent across sub-aperture images,

Weight-Space Geometry of Offline Reasoning Training

ResearchDGX agent

arXiv:2606.23740v1 Announce Type: cross Abstract: Offline reinforcement-learning losses (RFT, RIFT, DFT, Offline GRPO, DPO) are widely used to distill reasoning from large teachers into smaller studen

What Do Flow-Based Inverse Solvers Approximate? A Posterior-Transport View

SafetyDGX agent

arXiv:2606.24516v1 Announce Type: new Abstract: A growing family of training-free solvers -- FlowDPS, FLOWER, PnP-Flow and their diffusion ancestors (DPS, DAPS) -- repurpose a pretrained flow-matching

What Does ODRL Mean? A Cross-Level Ontological Grounding of Permissions, Prohibitions, and Duties in UFO-L

Model ReleasesDGX agent

arXiv:2606.24344v1 Announce Type: cross Abstract: ODRL policy evaluators produce verdicts, but say nothing about the normative positions a policy brings into existence, the authority structures those

What's Missing in Vision-Language Models? Probing Their Struggles with Causal Order Reasoning

ResearchDGX agent

arXiv:2506.00869v3 Announce Type: replace Abstract: Despite the impressive performance of vision-language models (VLMs) on downstream tasks, their ability to understand and reason about causal relatio

When AI Meets Finance (StockAgent): Large Language Model-based Stock Trading in Simulated Real-world Environments

SafetyDGX agent

arXiv:2407.18957v5 Announce Type: replace-cross Abstract: Can AI Agents simulate real-world trading environments to investigate the impact of external factors on stock trading activities (e.g., macroe

When CQs Go Wrong: Challenges in CQ Verification with OE-Assist

SafetyDGX agent

arXiv:2606.24619v1 Announce Type: new Abstract: Competency Questions (CQs) are the central component of CQ-verification, an established process in which an ontology is evaluated against a set of natur

When Helpfulness Overrides Causal Caution: Context-Dependent Suppression and Recovery in LLMs

Model ReleasesDGX agent

arXiv:2606.24370v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into decision-support roles in business and policy contexts. While prior benchmark studies have

When Preferences Fail to Become Incentives: A Utility-Behavior Gap in Large Language Models

SafetyDGX agent

arXiv:2606.22974v2 Announce Type: replace Abstract: Recent work on preference elicitation in large language models (LLMs) has demonstrated that, when given a series of choices between two outcomes, LL

When Retrieval Metrics Mislead: Measuring Policy Signal in Long-Horizon Tool-Use Agents

Model ReleasesDGX agent

arXiv:2606.23937v1 Announce Type: cross Abstract: Exact-match retrieval recall is often used as a proxy for whether a retriever supplies useful policy context to a downstream decision model. We test t

When Top-1 Fails: Calibrating LoRA Monitors for Masked Diffusion LMs

Model ReleasesDGX agent

arXiv:2606.24119v1 Announce Type: cross Abstract: Discrete diffusion language model (DLM) fine-tuning inherits inexpensive diagnostics from denoising-time confidence monitors, but their PEFT-training

Which Spaces can be Embedded in L_p-type Reproducing Kernel Banach Space? A Characterization via Metric Entropy

ResearchDGX agent

arXiv:2410.11116v4 Announce Type: replace-cross Abstract: In this paper, we establish a novel connection between the metric entropy growth and the embeddability of function spaces into reproducing ker

WiFi-Based People Counting Using Beam-Steerable Antennas: A Test-bed Study

ApplicationsDGX agent

arXiv:2606.23710v1 Announce Type: cross Abstract: Ubiquitous perception through RF signals is a pivotal opportunity for future technology: it enables personalized services such as smart living, remote

World Models in Pieces: Structural Certification for General Agents

Local AiDGX agent

arXiv:2606.24842v1 Announce Type: new Abstract: In the big-world regime, agents cannot be universally capable and their ability is inevitably specialized across a world model in pieces. Consequently,

World Value Models for Robotic Manipulation

Model ReleasesDGX agent

arXiv:2606.24742v1 Announce Type: new Abstract: Generalist value models play a pivotal role in scaling robotic policy learning from large-scale, mixed-quality data. Mathematically, accurate value esti

You Don't Need to Run Every Eval

Model ReleasesDGX agent

arXiv:2606.24020v1 Announce Type: new Abstract: A modern model release reports scores on 40+ benchmarks and the same evaluations were run many more times before it: to track training progress, compare

Zero-Shot Neural Priors for Generalizable Cross-Subject and Cross-Task EEG Decoding

ResearchDGX agent

arXiv:2606.23706v1 Announce Type: cross Abstract: The development of generalizable electroencephalography (EEG) decoding models is essential for robust brain-computer interfaces (BCI) and objective ne

Zero-Shot Test-Time Canonicalization using Out-of-Distribution Scoring

ResearchDGX agent

arXiv:2606.24178v1 Announce Type: cross Abstract: Pretrained vision models often misclassify inputs that are rotated, scaled, or sheared, even though these affine transformations leave the object clas

ZONOS2 Technical Report

Model ReleasesDGX agent

arXiv:2606.24320v1 Announce Type: cross Abstract: We present ZONOS2 8B, our latest TTS model, which achieves state-of-the-art naturalness, prosody, and voice cloning fidelity. We improve upon Zonos-v0

23 Jun 2026

2D Versus 3D Diffusion for In Silico Training of Interventional X-ray AI Models

ResearchDGX agent

arXiv:2606.21414v1 Announce Type: cross Abstract: The ability to synthesize realistic X-ray images has catalyzed the development of AI models for X-ray image-guided procedures, which otherwise suffer

360Anything: Geometry-Free Lifting of Images and Videos to 360{eg}

SafetyDGX agent

arXiv:2601.16192v2 Announce Type: replace Abstract: Lifting perspective images and videos to 360{eg} panoramas enables immersive 3D world generation. Existing approaches often rely on explicit geometr

3D Vessel Reconstruction from Sparse-View Dynamic DSA Images via Vessel Probability Guided Attenuation Learning

AgentsDGX agent

arXiv:2405.10705v3 Announce Type: replace-cross Abstract: Digital Subtraction Angiography (DSA) is one of the gold standards for vascular disease diagnosis. With the help of a contrast agent, time-res

4DVLT: Dynamic Scene Understanding with Worldline-Centered Vision-Language Tracking

Model ReleasesDGX agent

arXiv:2606.22631v1 Announce Type: new Abstract: 4D dynamic scene understanding requires grounding language to a persistent worldline that binds identity, metric 3D motion, and synchronized multi-view

A Causal DAG Prior for Synthetic Time-Series Classification Datasets

ResearchDGX agent

arXiv:2606.21776v1 Announce Type: new Abstract: A Prior-data fitted Network learns the posterior predictive induced by its training prior; bringing this paradigm to multivariate time-series classifica

A Completion-Aware Framework for Impactful Counterfactual Explainability in Graph Neural Networks

ApplicationsDGX agent

arXiv:2606.22033v1 Announce Type: new Abstract: In this study, we propose a novel pipeline for generic, model-agnostic, local-level counterfactual explainability in graph neural networks (GNNs). Altho

A Comprehensive Study on Visual Token Redundancy for Discrete Diffusion-based Multimodal Large Language Models

ResearchDGX agent

arXiv:2511.15098v2 Announce Type: replace Abstract: Discrete diffusion-based multimodal large language models (dMLLMs) have emerged as a promising alternative to autoregressive MLLMs thanks to their a

A Controlled Study of CLIP-Based Body-Scene Fusion for Emotion Recognition in Context

SafetyDGX agent

arXiv:2606.22072v1 Announce Type: new Abstract: Apparent emotion in natural images is often not visible from the face alone. The face may be small, hidden, or neutral, while posture and scene context

A Differentiable Atari VCS:A Complex, Fully Known Ground Truth for Explainable AI

HardwareDGX agent

arXiv:2606.22447v1 Announce Type: cross Abstract: Explanation requires ground truth: to verify an account of a system we must know its inner functioning-just what is missing where explainable AI (XAI)

A Digital Twin Framework for Traffic-Aware UAV Pavement Monitoring without Lane Closure

AgentsDGX agent

arXiv:2606.20742v1 Announce Type: new Abstract: UAV-based pavement inspection can reduce the cost and risk of road-surface monitoring, but real-world deployment remains difficult when traffic, pedestr

A DVDrive Approach for doScenes Instructed Driving Challenge

Local AiDGX agent

arXiv:2606.21623v1 Announce Type: new Abstract: Instruction-conditioned trajectory prediction is an emerging problem in autonomous driving, where a model predicts the future ego trajectory not only fr

A-Evolve-Training: Autonomous Post-Training of a 30B Model

Model ReleasesDGX agent

arXiv:2606.20657v1 Announce Type: cross Abstract: Post-training a frontier model is normally weeks of human work: proposing data and recipe changes, launching runs, reading evals, deciding what to kee

A Framework for Directed Acyclic Hypergraph Learning

ResearchDGX agent

arXiv:2606.21668v1 Announce Type: new Abstract: Continuous optimization methods for learning Directed Acyclic Graphs (DAGs) operate on weighted adjacency matrices and are therefore limited to pairwise

A Gated Graph Neural Network Approach to Fast-Convergent Dynamic Average Estimation

Local AiDGX agent

arXiv:2606.20955v1 Announce Type: new Abstract: Dynamic average estimation is a critical problem in multi-agent systems, enabling agents to collaboratively estimate time-varying signals using only loc

A Generalization Bound for Nearly-Linear Networks

ResearchDGX agent

arXiv:2407.06765v2 Announce Type: replace Abstract: We consider nonlinear networks as perturbations of linear ones. Based on this approach, we present novel generalization bounds that become non-vacuo

A Generalized Formalism of Auto-Regressive Decoding for Speech Processing

ResearchDGX agent

arXiv:2606.20714v1 Announce Type: cross Abstract: In speech processing, most state-of-the-art sequence prediction models rely on auto-regressive (AR) strategies to generate output sequences based on t

A Generative Model for Closed-Loop Microsimulation of Signalized Intersections

SafetyDGX agent

arXiv:2606.23588v1 Announce Type: new Abstract: Traffic microsimulators rely on hand-crafted behavior models that reproduce aggregate flow but miss the heterogeneous interactions between vehicles at s

A geometric and deep learning reproducible pipeline for monitoring floating anthropogenic debris in urban rivers using in situ cameras

ResearchDGX agent

arXiv:2510.23798v2 Announce Type: cross Abstract: The proliferation of floating anthropogenic debris in rivers has emerged as a pressing environmental concern, exerting a detrimental influence on biod

A Geometric Task-Space Port-Hamiltonian Formulation for Redundant Manipulators

ResearchDGX agent

arXiv:2512.14349v2 Announce Type: replace-cross Abstract: We present a novel geometric port-Hamiltonian formulation of redundant manipulators performing a differential kinematic task eta=J(q)ot{q}, wh

A Human-Inspired Thumb-Index Robotic Hand with Strain Gauges Embedded in Soft Joints

ResearchDGX agent

arXiv:2606.21245v1 Announce Type: new Abstract: Human hand grasp adaptation depends mainly on the synergy between physical structure and biological feedback. Inspired by this biomechanical principle,

A Hybrid, Multi-Layered Pipeline for Phishing and Threat Classification: Independently Validated URL and NLP Engines with a Calibrated Multi-Channel Fusion Stage

Model ReleasesDGX agent

arXiv:2606.21690v1 Announce Type: cross Abstract: Phishing is a multi-modal threat. We present a hybrid pipeline that scores each modality with its own engine and fuses the results. Three engines are

A Hybrid TGN-SEAL Model for Dynamic Graph Link Prediction

TutorialsDGX agent

arXiv:2602.14239v2 Announce Type: replace-cross Abstract: Predicting links in sparse, continuously evolving networks is a central challenge in network science. Conventional heuristic methods and deep

A Jointly Efficient and Optimal Algorithm for Heteroskedastic Generalized Linear Bandits with Adversarial Corruptions

ResearchDGX agent

arXiv:2602.10971v2 Announce Type: replace Abstract: We consider the problem of heteroskedastic generalized linear bandits (GLBs) with adversarial corruptions, which subsumes heteroskedastic linear ban

A large-scale foundation model enables simulation-to-real adaptation for nuclear magnetic resonance-based molecular structure analysis

TutorialsDGX agent

arXiv:2606.20756v1 Announce Type: cross Abstract: Nuclear Magnetic Resonance (NMR) spectroscopy is a powerful tool for molecular structure analysis, and spectral artificial intelligence offers great p

A Latent Representation Learning Framework for Hyperspectral Image Emulation in Remote Sensing

Model ReleasesDGX agent

arXiv:2603.21911v2 Announce Type: replace Abstract: Synthetic hyperspectral image (HSI) generation is essential for large-scale simulation, algorithm development, and mission design, yet traditional r

← Previous
1…382383384385386…1040
Next →