AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,617
  • Agents7,497
  • Applications5,365
  • Concepts5
  • Hardware1,816
  • Industry6,151
  • Local Ai4,900
  • Model Releases23,593
  • Research19,967
  • Safety13,267
  • Syntheses17
  • Tools1,674
  • Tutorials3,365

Source
Human
87,617Total entries
1Added by human
87,616Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
62,383 results
24 Jun 2026

Grounding Generative Policies in Physics: Optimization-Guided Diffusion for Robot Control

ResearchDGX agent

arXiv:2606.24208v1 Announce Type: new Abstract: Diffusion models sample effectively from high-dimensional, multimodal distributions, but their outputs may violate deployment constraints. For task-spac

GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents

Model ReleasesDGX agent

arXiv:2606.24551v1 Announce Type: new Abstract: Computer-use agents can execute software tasks through either graphical interfaces or programmatic command interfaces, but existing evaluations confound

HANCLIP: A Family of Hyperbolic Angular Negation Vision Language Models

Model ReleasesDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.23843v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are typically pre-trained on large-scale image-text datasets to capture semantic correspondences between visual content an

Hardware-Oriented Inference Complexity of Kolmogorov-Arnold Networks

HardwareDGX agent

arXiv:2604.03345v2 Announce Type: replace Abstract: Kolmogorov-Arnold Networks (KANs) have recently emerged as a powerful architecture for various machine learning applications. However, their unique

Harmonic: Hierarchical State Space Models for Efficient Long-Context Language Modeling

Model ReleasesDGX agent

arXiv:2606.24650v1 Announce Type: new Abstract: We present Harmonic, a hierarchical state space model (SSM) for language modeling. The architecture stacks three recurrent levels at progressively slowe

Hessian-augmented Supervised Learning for Hamilton-Jacobi-Bellman PDEs

ResearchDGX agent

arXiv:2606.23827v1 Announce Type: cross Abstract: A data-driven method is developed for approximating value functions in deterministic optimal control problems with nonlinear control-affine dynamics.

Heterogeneous 2D/1D Signal Representation Fusion for Underwater Acoustic Modulation Recognition Under Distribution Shift

Model ReleasesDGX agent

arXiv:2606.23702v1 Announce Type: cross Abstract: Modulation recognition systems rely on heterogeneous signal representations. 2D signal-image modalities such as time-frequency and cyclostationary map

Heterogeneous Knowledge Distillation via Geometry Decoupling and Momentum-Aware Gradient Regulation

Model ReleasesDGX agent

arXiv:2606.24557v1 Announce Type: new Abstract: Heterogeneous Knowledge Distillation (HKD) aims to transfer knowledge across varying architectures (e.g., from Transformer to CNN) but inherently suffer

Hierarchical Spatial and Channel Aggregation for Cross-domain Few-shot Segmentation

SafetyDGX agent

arXiv:2606.24296v1 Announce Type: new Abstract: Cross-domain Few-shot Segmentation (CD-FSS) aims to learn generalizable segmentation capability from abundant annotated samples in the source domain, en

High-Fidelity Synthetic Transmission Electron Microscopy Image Generation Using Diffusion Probabilistic Models for Data-Limited Semiconductor Metrology

ResearchDGX agent

arXiv:2606.24817v1 Announce Type: new Abstract: Advanced semiconductor nodes drastically increased demand for Transmission Electron Microscopy (TEM), yet destructive sample preparation, slow imaging a

HiPath: Hierarchical Vision-Language Alignment for Structured Pathology Report Prediction

SafetyDGX agent

arXiv:2603.19957v2 Announce Type: replace-cross Abstract: Pathology reports are structured, multi-granular documents encoding diagnostic conclusions, histological grades, and ancillary test results ac

Holistic Data Scheduler for LLM Pre-training via Multi-Objective Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.24133v1 Announce Type: cross Abstract: The composition of training data, governed by the diversity of sources and their mixing strategy, is a cornerstone of Large Language Model (LLM) pre-t

HOLMES: Evaluating Higher-Order Logical Reasoning in LLMs

Model ReleasesDGX agent

arXiv:2606.23238v2 Announce Type: replace Abstract: Logical reasoning is essential for reliable AI, yet existing benchmarks are largely first-order-logic-centric, focusing on object-level deduction ov

How Much Can We Trust LLM Search Agents? Measuring Endorsement Vulnerability to Web Content Manipulation

Model ReleasesDGX agent

arXiv:2606.16821v2 Announce Type: replace Abstract: Large language model (LLM)-based search agents synthesize open-web content into actionable recommendations on behalf of users, creating a risk that

Hybrid Event Frame Sensors: Modeling, Calibration, and Simulation

ResearchDGX agent

arXiv:2511.18037v2 Announce Type: replace Abstract: Hybrid event-frame sensors integrate an Event Vision Sensor (EVS) and an Active Pixel Sensor (APS) within a single chip, combining the high dynamic

Hybrid Sequence Modeling and Reinforced Verification for Controllable Target-Conditioned Decision Making

SafetyDGX agent

arXiv:2508.16420v3 Announce Type: replace Abstract: Target-conditioned sequence models provide a simple interface for controllable offline decision making, but the requested target return can be an un

HyMaTE: A Hybrid Mamba and Transformer Model for EHR Representation Learning

ApplicationsDGX agent

arXiv:2509.24118v2 Announce Type: replace Abstract: Electronic health Records (EHRs) have become a cornerstone in modern-day healthcare. They are a crucial part for analyzing the progression of patien

Ill-Posed by Design: Probing Evidence Use in VLMs

Local AiDGX agent

arXiv:2606.24335v1 Announce Type: new Abstract: Counterfactual analysis is widely used to study evidence use in vision-language models, but its diagnostic value is limited on well-posed tasks: when se

Impatient Bandits: Optimizing for the Long-Term Without Delay

ResearchDGX agent

arXiv:2501.07761v2 Announce Type: replace-cross Abstract: Increasingly, recommender systems are tasked with improving users' long-term satisfaction. In this context, we study a content exploration tas

Importance of Intent-Sharing for V2X-based Maneuver Coordination

ResearchDGX agent

arXiv:2606.24203v1 Announce Type: cross Abstract: This paper examines the critical role of intent-sharing in enabling effective maneuver coordination for connected and automated vehicles (CAVs). Succe

Inclusive Interactive Collisions for Multi-View Consistent Compositional 3D Generation

TutorialsDGX agent

arXiv:2606.24206v1 Announce Type: cross Abstract: Recent breakthroughs in 3D generation have advanced notably with the development of text-to-image diffusion model. However, existing methods remain tw

Infinitesimal Causality

ResearchDGX agent

arXiv:2606.24621v1 Announce Type: cross Abstract: This paper introduces a categorical account of infinitesimal causality in Frobenius Markov categories equipped with tangent-bundle semantics. IDC capt

Information-Theoretic Classifier-Free Guidance with Adaptive Schedule Optimization

ResearchDGX agent

arXiv:2606.24025v1 Announce Type: new Abstract: Diffusion models have achieved strong performance in image, text-to-image, and video generation, where conditional generation is often controlled by cla

Ingredient-Level Food Image Segmentation for Nutrition Awareness

ResearchDGX agent

arXiv:2606.24059v1 Announce Type: new Abstract: Food images often contain several visible ingredients, so assigning one dish label to an entire image hides important visual structure. This work studie

InSight: Self-Guided Skill Acquisition via Steerable VLAs

AgentsDGX agent

arXiv:2606.24884v1 Announce Type: cross Abstract: Vision-language-action (VLA) models can learn manipulation skills from demonstrations, but their capabilities are bounded by the skills in the trainin

Integrated Sensing and Communications for Real-time Avatar Control in XR over 5G

ResearchDGX agent

arXiv:2606.23771v1 Announce Type: cross Abstract: Extended Reality (XR) presents a challenging use case for 5G and 6G networks, requiring high data-rates and lowlatency communication to deliver a trul

Invariant Graph Representations for Continuous-Time Dynamic Graphs Under Distribution Shifts

TutorialsDGX agent

arXiv:2405.19062v2 Announce Type: replace-cross Abstract: Continuous-Time Dynamic Graphs (CTDGs) enable fine-grained modeling of evolving relational systems. However, most existing CTDG representation

IPO Finance Agent: Evaluation of LLM Financial Analysts beyond Finance Agent v2, with Automated Rubric Generation -- the Case of the SpaceX (SPCX) IPO

Model ReleasesDGX agent

arXiv:2606.23032v2 Announce Type: replace Abstract: Finance Agent v2 (by Vals AI) has emerged as the reference benchmark for evaluating both Anthropic Claude and OpenAI ChatGPT frontier language model

It's Complicated: On the Design and Evaluation of AI-Powered AAC Interfaces

ResearchDGX agent

arXiv:2606.24854v1 Announce Type: cross Abstract: Artificial intelligence (AI) can enhance what people who use augmentative and alternative communication (AAC) are able to do with their systems. Howev

IV-CoT: Implicit Visual Chain-of-Thought for Structure-Aware Text-to-Image Generation

TutorialsDGX agent

arXiv:2606.24849v1 Announce Type: cross Abstract: Unified multi-modal large language models (MLLMs) have achieved strong text-to-image generation quality, but still struggle with structure-aware promp

JEDEL: Zero-Shot DNA-Encoded Library Design for Early-Stage Drug Discovery

SafetyDGX agent

arXiv:2606.23745v1 Announce Type: cross Abstract: We present JEDEL, a framework for generating synthesis-ready DNA-encoded libraries (DELs) directly from three-dimensional pharmacophore representation

Jolia: Concept-Level Vision-Language Alignment for 3D CT Contrastive Learning

Model ReleasesDGX agent

arXiv:2606.24570v1 Announce Type: new Abstract: Vision-language contrastive pretraining has become the dominant recipe for 3D medical foundation models, leveraging the large volumes of paired scans an

JupOtter: Cell-Level Bug Detection in Jupyter Notebooks

ResearchDGX agent

arXiv:2606.23877v1 Announce Type: cross Abstract: Jupyter Notebooks are an increasingly popular coding environment used across many domains, especially in Python-based data science and scientific comp

KANLib -- A Modular, Extensible and Fast Kolmogorov-Arnold Network Implementation

Model ReleasesDGX agent

arXiv:2606.17927v2 Announce Type: replace-cross Abstract: Kolmogorov-Arnold Networks (KANs) have recently emerged as a promising alternative to traditional multilayer perceptrons by replacing linear w

KLip-PPO: A per-sample KL perspective on PPO-Clip

SafetyDGX agent

arXiv:2606.23932v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) is the standard policy-gradient algorithm for on-policy reinforcement learning. The literature presents it in two for

Knowledge-Graph Grounding Helps LLMs Only for Out-of-Training Knowledge: A Controlled Study on Clinical Question Answering

Model ReleasesDGX agent

arXiv:2606.22419v2 Announce Type: replace Abstract: A recent Nature Medicine study reports that general-purpose frontier LLMs outperform specialized retrieval-augmented clinical tools on medical bench

L3Cube-MahaPOS: A Marathi Part-of-Speech Tagging Dataset and BERT Models

Model ReleasesDGX agent

arXiv:2606.24825v1 Announce Type: new Abstract: Part-of-Speech (POS) tagging is a foundational NLP task underpinning machine translation, information extraction, and syntactic parsing. Despite Marathi

LaGO: Latent Action Guidance for Online Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.24669v1 Announce Type: new Abstract: Large language models (LLMs) have shown strong potential for planning and sequential decision-making, but prior work often relies on using them as direc

LangMAP: A Language-Adaptive Approach to Tokenization

Model ReleasesDGX agent

arXiv:2606.23566v2 Announce Type: replace Abstract: Language-specific tokenizers improve tokenization quality and the downstream performance of models on those languages. However, using such a tokeniz

Large-Language-Model Discovery of Quantum LDPC Codes through Structured Concept Evolution

Model ReleasesDGX agent

arXiv:2606.24808v1 Announce Type: cross Abstract: Quantum computers could outperform classical machines on important problems, but only if the errors that pervade quantum hardware can be corrected at

Latent Visual States for Efficient Multimodal Reasoning

SafetyDGX agent

arXiv:2606.24233v1 Announce Type: new Abstract: The integration of visual evidence has significantly enhanced the capabilities of large multimodal models. However, this integration predominantly relie

Layer-wise Probing of wav2vec 2.0 and Whisper for Consonant Cluster Reduction in African American English

ResearchDGX agent

arXiv:2606.23948v1 Announce Type: new Abstract: Self-supervised and supervised speech models are increasingly used to investigate which linguistic information their internal representations encode, an

Learning the Koopman Operator using Attention Free Transformers

Model ReleasesDGX agent

arXiv:2606.23957v1 Announce Type: new Abstract: Learning Koopman operators with autoencoders enables linear prediction in a latent space, but long-horizon rollouts often drift off the learned manifold

Learning to Trigger: Reinforcement Learning at the Large Hadron Collider

Model ReleasesDGX agent

arXiv:2606.23993v1 Announce Type: cross Abstract: High-throughput scientific facilities such as the Large Hadron Collider depend on real-time event filtering (extit{triggering}) under tight constraint

LecturaAgents: A Multi-Agent Framework for Adaptive Personalized AI-Assisted Learning and Embodied Teaching

SafetyDGX agent

arXiv:2606.16428v2 Announce Type: replace-cross Abstract: Effective personalized AI-assisted learning demands systems that can not only generate accurate learner-specific educational materials, but al

Legal Reasoning Is Not Lawyering: Rethinking Legal Benchmarks for Pro Se Access to Justice

Model ReleasesDGX agent

arXiv:2606.23716v1 Announce Type: cross Abstract: Legal AI benchmark research frequently invokes the assumption that large language models can improve access to justice, including for people who canno

Legible and Intuitive Multi-modal Robot State and Intent Communication Validated in Online and Real-world Studies

ApplicationsDGX agent

arXiv:2606.24445v1 Announce Type: new Abstract: Effective robot-to-human communication can increase transparency and trust, reduce uncertainty, and contribute to safer collaboration in shared workspac

LemonHarness Technical Report

Model ReleasesDGX agent

arXiv:2606.24311v1 Announce Type: new Abstract: As large language model (LLM) agents are applied to longer tasks, they increasingly modify workspace state across multiple rounds of iteration. However,

Less is More: Quality-Aware Training Data Selection for Scientific Summarization

SafetyDGX agent

arXiv:2606.24828v1 Announce Type: new Abstract: Scientific long-document summarization datasets commonly treat author-written abstracts as gold reference summaries, although their quality and alignmen

Light-weight Pronunciation Assessment via Discrete Speech Token Surprisal

SafetyDGX agent

arXiv:2606.19910v2 Announce Type: replace Abstract: Training automated pronunciation assessment often relies on labeled learner errors or non-native corpora that are costly to collect. We propose a li

Lightweight Test-Time Adaptation for EMG-Based Gesture Recognition

SafetyDGX agent

arXiv:2601.04181v2 Announce Type: replace Abstract: Reliable long-term decoding of gestures from surface electromyography (EMG) is hindered by signal drift caused by electrode displacement, muscle fat

Lightweight Transformer Models for On-Device Fault Detection: A Benchmark Study on Resource-Constrained Deployment

Model ReleasesDGX agent

arXiv:2606.24173v1 Announce Type: cross Abstract: On-device fault detection enables real-time diagnostics without cloud dependency, but deploying machine learning models on resource-constrained hardwa

Listening makes Vision Clear for VLMs

SafetyDGX agent

arXiv:2606.23763v1 Announce Type: cross Abstract: Recent work typically assesses vision--language consistency using attention distributions of answer-side tokens. However, we observe that highest atte

Lite Any Stereo V2: Faster and Stronger Efficient Zero-Shot Stereo Matching

ApplicationsDGX agent

arXiv:2606.24457v1 Announce Type: new Abstract: Recent advances in stereo matching have achieved remarkable accuracy, but often rely on large models, heavy computation, or additional foundation-model

LLM-MINE: Large Language Model based Alzheimer's Disease and Related Dementias Phenotypes Mining from Clinical Notes

ResearchDGX agent

arXiv:2603.13673v2 Announce Type: replace Abstract: Accurate extraction of Alzheimer's Disease and Related Dementias (ADRD) phenotypes from electronic health records (EHR) is critical for early-stage

LLMs are Bayesian, In Expectation, Not in Realization

ResearchDGX agent

arXiv:2507.11768v3 Announce Type: replace-cross Abstract: Bayesian accounts of in-context learning face a direct objection: exact posterior predictives for exchangeable data are invariant to task-pres

LLMs Prompted for Legal Context Object More: Overrefusal from Small On-Premises LLMs in Criminal Legal Context

Local AiDGX agent

arXiv:2606.24585v1 Announce Type: new Abstract: While the validity of LLMs' use in the legal context remains subject to ethical and legal debate, legal professionals are already experimenting with per

LoMime: Query-Efficient Membership Inference using Model Extraction in Label-Only Settings

Model ReleasesDGX agent

arXiv:2602.18934v2 Announce Type: replace Abstract: Membership inference attacks (MIAs) threaten the privacy of machine learning models by revealing whether a specific data point was used during train

Loss Landscape Poisoning: Targeted Extraction of Unseen Training Data from LLMs

Local AiDGX agent

arXiv:2606.17110v2 Announce Type: replace-cross Abstract: Large Language Models are increasingly trained on proprietary or sensitive data, from private healthcare and financial records to user convers

LoT-Pass: Long-term-robust Image Watermarking for Image to Video Generation

Model ReleasesDGX agent

arXiv:2509.17773v2 Announce Type: replace Abstract: The rapid progress of image-guided video generation (I2V) has raised concerns about its potential misuse in misinformation and fraud, underscoring t

← Previous
1…377378379380381…1040
Next →