AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,620 results
2 Jun 2026

MOC: Multi-Order Communication in LLM-based Multi-Agent Systems

Model ReleasesDGX agent

arXiv:2606.02359v1 Announce Type: new Abstract: Despite the remarkable progress of Large Language Model (LLM) based Multi-Agent Systems, most research focuses on optimizing coordination topology while

Model-Based Quality Assessment for Massively Multilingual Parallel Data

Model ReleasesDGX agent

arXiv:2606.00285v1 Announce Type: new Abstract: Large-scale multilingual bitext often contains two distinct problems: non-parallel sentence pairs and low-quality translations. We decompose model-based

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

Model Parallelism With Subnetwork Data Parallelism

Model ReleasesDGX agent

arXiv:2507.09029v5 Announce Type: replace-cross Abstract: Pre-training large neural networks at scale imposes heavy memory demands on accelerators and often requires costly communication. We introduce

Molecular Embedding-Based Algorithm Selection in Protein-Ligand Docking

Model ReleasesDGX agent

arXiv:2512.02328v2 Announce Type: replace-cross Abstract: Selecting an effective docking algorithm is highly context-dependent, and no single method performs reliably across structural, chemical, and

Moment-Video: Diagnosing Temporal Fidelity of Video MLLMs on Momentary Visual Events

Model ReleasesDGX agent

arXiv:2606.02522v1 Announce Type: cross Abstract: Video multimodal large language models (MLLMs) have made rapid progress on general and long-form video understanding, yet their ability to preserve br

MomentKV: Closing the Directional Gap in KV Cache Eviction for Long-Context Inference

Model ReleasesDGX agent

arXiv:2606.01563v1 Announce Type: new Abstract: Autoregressive decoding in Transformer-based language models relies on the KV cache, whose memory footprint grows linearly with sequence length and beco

Momento: Evaluating Persistent Memory and Reasoning with Multi-Session Agentic Conversations

Model ReleasesDGX agent

arXiv:2606.00832v1 Announce Type: new Abstract: Recent advances in agentic AI have enabled agents to complete complex tasks through tool use, reasoning, and multi-step planning. Yet existing benchmark

Motion-aware Event Suppression for Event Cameras

Model ReleasesDGX agent

arXiv:2602.23204v3 Announce Type: replace Abstract: In this work, we introduce the first framework for Motion-aware Event Suppression, which learns to filter events triggered by IMOs and ego-motion in

MotionDreamer: Universal Skeletal Motion Generation for 3D Rigged Shapes

Model ReleasesDGX agent

arXiv:2606.01518v1 Announce Type: new Abstract: Motion generation for rigged shapes is vital for scalable 4D asset production. However, template-based methods are limited by specific topologies and fa

Move the Query, Not the Cache: Characterizing Cross-Instance Latent Attention Redistribution Across GPU Fabrics

Model ReleasesDGX agent

arXiv:2606.01502v1 Announce Type: cross Abstract: Frontier LLMs increasingly decide what a query attends to with a sparse-attention indexer that picks a few KV-cache blocks per query: attention's unit

MT-EditFlow: Reinforcement Learning for Multi-Turn Image Editing with Flow Matching

Model ReleasesDGX agent

arXiv:2606.01985v1 Announce Type: new Abstract: Recent breakthroughs in instruction-based image editing have captured significant attention, as models are now capable of handling real-world editing de

Multi-Agent Computer Use

Model ReleasesDGX agent

arXiv:2606.01533v1 Announce Type: cross Abstract: Computer use agents (CUAs) today are primarily deployed as single serial agents. This setup is suboptimal for complex long-horizon tasks that benefit

Multi-Contrast MRI Motion Correction via Parameter-Informed Disentanglement and Adaptive Experts

Model ReleasesDGX agent

arXiv:2606.00146v1 Announce Type: cross Abstract: Motion artifacts in magnetic resonance imaging (MRI) degrade diagnostic reliability. Existing deep learning methods are typically contrast-specific an

Multimodal Action Diffusion for Robust End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.02105v1 Announce Type: new Abstract: End-to-End Autonomous Driving (E2E-AD) systems have largely converged on predicting intermediate trajectory waypoints, delegating final control to hand-

Multimodal Approaches for Visually-Rich Document Type Classification: A Comparative Analysis

Model ReleasesDGX agent

arXiv:2606.02162v1 Announce Type: cross Abstract: Document type classification in visually rich documents remains challenging, as relevant information is distributed across textual, visual, and layout

Multimodal Music Recommendation System using LLMs

Model ReleasesDGX agent

arXiv:2606.00125v1 Announce Type: cross Abstract: Music recommendation systems typically treat songs as opaque tokens, relying on collaborative interaction histories which overlooks semantic or acoust

naPINN: Noise-Adaptive Physics-Informed Neural Networks for Recovering Physics from Corrupted Measurement

Model ReleasesDGX agent

arXiv:2602.02547v2 Announce Type: replace-cross Abstract: Physics-Informed Neural Networks (PINNs) are effective methods for solving inverse problems and discovering governing equations from observati

Navigating the Reality Gap: On-Device Continual Adaptation of ASR for Clinical Telephony

Model ReleasesDGX agent

arXiv:2512.16401v5 Announce Type: replace Abstract: Automatic Speech Recognition (ASR) can significantly reduce documentation burden in clinical workflows, but standard models degrade sharply in real-

Neural Network Compression by Approximate Differential Equivalence

Model ReleasesDGX agent

arXiv:2606.01402v1 Announce Type: cross Abstract: Neural network compression is commonly achieved by pruning parameters based on local importance scores, e.g., magnitude-based pruning. We propose a co

new in deepagents: agent rubrics! you define a rubric, and the agent self-evaluates and iterates until it satisfies every rubric criterion. …

Model ReleasesDGX agent

new in deepagents: agent rubrics! you define a rubric, and the agent self-evaluates and iterates until it satisfies every rubric criterion. this is similar to /goal in claude code or codex, but more f

NILC: Discovering New Intents with LLM-assisted Clustering

Model ReleasesDGX agent

arXiv:2511.05913v2 Announce Type: replace-cross Abstract: New intent discovery (NID) seeks to recognize both new and known intents from unlabeled user utterances, which finds prevalent use in practica

Non-Learning Low-Light Stereo Vision

Model ReleasesDGX agent

arXiv:2606.00379v1 Announce Type: new Abstract: We present a non-learning stereo framework for disparity estimation from severely noisy images. Using the Field of Junctions (FoJ), it retains coarse vi

NOS-Gate: Queue-Aware Streaming IDS for Consumer Gateways under Timing-Controlled Evasion

Model ReleasesDGX agent

arXiv:2601.00389v2 Announce Type: replace-cross Abstract: Timing and burst patterns can leak through encryption, and an adaptive adversary can exploit them. This undermines metadata-only detection in

OARelatedWork: A Large-Scale Dataset of Related Work Sections with Full-texts from Open Access Sources

Model ReleasesDGX agent

arXiv:2405.01930v2 Announce Type: replace Abstract: This paper introduces OARelatedWork: a dataset for related work generation from open-access sources. It is the first large-scale multi-document summ

Off-the-Shelf LLMs as Process Scorers: Training-Free Alternative to PRMs for Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2606.01682v1 Announce Type: cross Abstract: Selecting the best response from multiple small-model samples using a stronger scorer is a simple inference-time strategy, but fails when the small mo

OmniEEG-Bench: A Standardized Evaluation Benchmark for EEG Foundation Models

Model ReleasesDGX agent

arXiv:2606.00815v1 Announce Type: new Abstract: Electroencephalography (EEG) supports a variety of brain-computer interface (BCI) tasks ranging from brain-state monitoring to human-LLM interactions. E

OmniOPD: Logit-Free On-Policy Distillation via Speculative Verification

Model ReleasesDGX agent

arXiv:2606.01476v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) trains a student model on its own generative trajectories under dense token-level feedback from a stronger teacher, mitig

On the Evaluation of Spiking Neural Network Configurations for Network Intrusion Detection

Model ReleasesDGX agent

arXiv:2606.01442v1 Announce Type: cross Abstract: Network intrusion detection is a core component of modern cybersecurity infrastructure, yet the deep learning models that dominate the field are compu

On the Generalization Gap in Self-Evolving Language Model Reasoning

Model ReleasesDGX agent

arXiv:2606.01075v1 Announce Type: new Abstract: Recent work suggests that large language models (LLMs) can improve through self-evolution (SE), using supervision signals generated by the model itself.

On the Generalization in Topology Optimization via Sensitivity-Conditioned Bernoulli Flow Matching

Model ReleasesDGX agent

arXiv:2606.02179v1 Announce Type: cross Abstract: Surrogate models for topology optimization (TO) exhibit highly variable out-of-distribution (OOD) generalization under distribution shifts such as cha

On the Limits of Token Reduction for Efficient Unified Vision Language Training

Model ReleasesDGX agent

arXiv:2606.01503v1 Announce Type: cross Abstract: Unified vision-language models (VLMs) integrate visual understanding and visual generation within a single autoregressive backbone, but their joint tr

On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters

Model ReleasesDGX agent

arXiv:2606.02437v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) is usually treated as a cheaper alternative to full fine-tuning. We study a broader role: small trainable adapt

On the Theoretical Limitations of Embedding-based Link Prediction

Model ReleasesDGX agent

arXiv:2506.22271v3 Announce Type: replace Abstract: Neural networks often map low-dimensional embeddings to high-dimensional output spaces. Usually, the output layer is linear, which can create a 'ran

On Wednesdays, We Ask Questions: Optimizing 'Active Listening' in Automated Legal Triage and Referral

Model ReleasesDGX agent

arXiv:2606.00272v1 Announce Type: new Abstract: The FETCH classifier generates follow-up questions to help refine the best match for the applicant's legal problem, using a low-cost ensemble of LLMs. I

OncoReason: Structuring Clinical Reasoning in LLMs for Robust and Interpretable Survival Prediction

Model ReleasesDGX agent

arXiv:2510.17532v2 Announce Type: replace Abstract: Predicting cancer treatment outcomes requires models that are both accurate and interpretable, particularly in the presence of heterogeneous clinica

OneVLA: A Unified Framework for Embodied Tasks

Model ReleasesDGX agent

arXiv:2606.01241v1 Announce Type: new Abstract: Navigation and manipulation are fundamental capabilities of embodied intelligence, enabling robots to interpret natural language commands and interact p

Ontology-Guided Reasoning for Affordance-Based Explanations of Robot Navigation

Model ReleasesDGX agent

arXiv:2606.00117v1 Announce Type: new Abstract: This paper proposes ontology-guided reasoning for affordance-based explanations of robot navigation. In human environments, it is not sufficient for a r

OpenAI extends Codex with productivity tools for nontechnical users

Model ReleasesDGX agent

OpenAI Group PBC today released a set of features that will make it easier for nontechnical people to use its Codex automation tool. The update comes five months after Anthropic PBC added similar capa

OpenAI releases a new report on knowledge work: Codex now has 5M+ weekly active users, up 6x+ since February, and knowledge workers are ~20% of Codex users (OpenAI)

Model ReleasesDGX agent

OpenAI: OpenAI releases a new report on knowledge work: Codex now has 5M+ weekly active users, up 6x+ since February, and knowledge workers are ~20% of Codex users — OpenAI today released a new report

OpenDPR: Open-Vocabulary Change Detection via Vision-Centric Diffusion-Guided Prototype Retrieval for Remote Sensing Imagery

Model ReleasesDGX agent

arXiv:2603.27645v2 Announce Type: replace Abstract: Open-vocabulary change detection (OVCD) seeks to recognize arbitrary changes of interest by enabling generalization beyond a fixed set of predefined

OpenWebRL: Demystifying Online Multi-turn Reinforcement Learning for Visual Web Agents

Model ReleasesDGX agent

arXiv:2606.02031v1 Announce Type: cross Abstract: Building capable visual web agents requires long-horizon reasoning, precise grounding, and robust interaction with dynamic real-world websites. Despit

Optimal Regularization for Performative Learning

Model ReleasesDGX agent

arXiv:2510.12249v2 Announce Type: replace Abstract: In performative learning, the data distribution reacts to the deployed model - for example, because strategic users adapt their features to game it

Order within Chaos: Capturing Intrinsic Energy Anomalies for AI-Manipulated Image Forgery Localization

Model ReleasesDGX agent

arXiv:2606.02178v1 Announce Type: cross Abstract: Recent advancements in generative AI have led to image editing models capable of producing realistic forgeries that evade traditional image forgery lo

PaintBench: Deterministic Evaluation of Precise Visual Editing

Model ReleasesDGX agent

arXiv:2606.00188v1 Announce Type: cross Abstract: While current multimodal models are proficient at open-ended visual editing, executing precise single-answer edits remains an important obstacle. To p

PaperVoyager : Building Interactive Web with Visual Language Models

Model ReleasesDGX agent

arXiv:2603.22999v3 Announce Type: replace Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, exist

Parameter Alignment Mitigates Catastrophic Forgetting in Multilingual Expert Language Models

Model ReleasesDGX agent

arXiv:2606.00284v1 Announce Type: new Abstract: While continual pretraining~(CPT) is a practical way to extend large language models to new languages, naive finetuning on targeted data erodes existing

Parameter-efficient Dual-encoder Architecture with Differentiable Choquet Integral Fusion for Underwater Acoustic Classification

Model ReleasesDGX agent

arXiv:2606.02341v1 Announce Type: cross Abstract: Underwater acoustic classification has a wide array of oceanic applications, but faces challenges due to an increasingly complex acoustic environment.

Parameter-Efficient Fine-Tuning of Large Pretrained Models for Instance Segmentation Tasks

Model ReleasesDGX agent

arXiv:2606.01947v1 Announce Type: cross Abstract: Research and applications in artificial intelligence have recently shifted with the rise of large pretrained models, which deliver state-of-the-art re

Parameter-Free and Group Conditional Online Conformal Prediction

Model ReleasesDGX agent

arXiv:2606.00419v1 Announce Type: cross Abstract: Uncertainty quantification (UQ) is critical for the deployment of machine learning predictors in real-world scenarios where the data distribution may

PaSBench-Video: A Streaming Video Benchmark for Proactive Safety Warning

Model ReleasesDGX agent

arXiv:2606.02443v1 Announce Type: cross Abstract: Between the first visible sign of danger and the moment an accident occurs, there is often a window where intervention remains possible. Video-capable

Pasted File Editor

Model ReleasesDGX agent

Tool: Pasted File Editor I really like how you can paste a large volume of text into claude.ai (or the Claude desktop/mobile apps) and it will detect it as a large paste and turn it into a file attach

Pause and Think: A Dataset and Benchmark for Video-Grounded Assistive Action Suggestion

Model ReleasesDGX agent

arXiv:2606.00616v1 Announce Type: cross Abstract: Recent Vision-Language Models (VLMs) struggle with grounded reasoning, temporal consistency, and context aware planning in videos. We introduce pause-

People who say this kind of thing are completely lost about what I actually said about deep learning, and I would strongly encourage them to…

Model ReleasesDGX agent

People who say this kind of thing are completely lost about what I actually said about deep learning, and I would strongly encourage them to read “Deep learning is hitting a wall” (2022). What I said

Perception First: A Frontier Native-Video Model with Self-Consistency for Implicit Video Question Answering

Model ReleasesDGX agent

arXiv:2606.01485v1 Announce Type: new Abstract: We describe our submission to the VRR Challenge @ CVPR 2026, built on the ImplicitQA / VRR-QA benchmark~ite{implicitqa}: multiple-choice video question

Permissive Safety Through Trusted Inference: Verifiable Belief-Space Neural Safety Filters for Assured Interactive Robotics

Model ReleasesDGX agent

arXiv:2606.02562v1 Announce Type: cross Abstract: Autonomous robots that interact with people must make safe and efficient decisions under human-induced uncertainty, such as their preferences, goals,

Persona Attack: Incremental Memory Injection Jailbreak Attack against Large Language Models

Model ReleasesDGX agent

arXiv:2606.00150v1 Announce Type: cross Abstract: As Large Language Models evolve for user convenience, vulnerability to jailbreak attacks continues to be reported despite ongoing efforts in safety tr

Personalized 3D Myocardial Infarct Geometry Reconstruction from Cine MRI for Cardiac Digital Twins

Model ReleasesDGX agent

arXiv:2606.01808v1 Announce Type: new Abstract: Accurate 3D geometric characterization of myocardial infarction (MI) is essential for building cardiac digital twins (CDTs) to precisely simulate infarc

Perturbative methods for non-parametric instrumental variable

Model ReleasesDGX agent

arXiv:2606.00322v1 Announce Type: new Abstract: We introduce a perturbative approach for nonparametric instrumental variable (NPIV) estimation. By drawing inspiration from perturbation theory in physi

PFT: Phonon Fine-tuning for Machine Learned Interatomic Potentials

Model ReleasesDGX agent

arXiv:2601.07742v4 Announce Type: replace-cross Abstract: Many materials properties depend on higher-order derivatives of the potential energy surface, yet machine learned interatomic potentials (MLIP

← Previous
1…182183184185186…377
Next →