AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,403
  • Agents7,557
  • Applications5,411
  • Concepts5
  • Hardware1,836
  • Industry6,170
  • Local Ai4,931
  • Model Releases23,900
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,403Total entries
1Added by human
88,402Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,624 results
19 May 2026

Symmetry-Compatible Principle for Optimizer Design: Embeddings, LM Heads, SwiGLU MLPs, and MoE Routers

Model ReleasesDGX agent

arXiv:2605.18106v1 Announce Type: cross Abstract: A striking geometric disparity has long persisted in the practice of deep learning. While modern neural network architectures naturally exhibit rich s

Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms

Model ReleasesDGX agent

arXiv:2605.17336v1 Announce Type: cross Abstract: Tactile sensing is a fundamental modality for embodied intelligence, offering unique and direct feedback on contact geometry, material properties, and

Taxonomy and Consistency Analysis of Safety Benchmarks for AI Agents

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.16282v1 Announce Type: cross Abstract: The rapid deployment of LLM-based autonomous agents has introduced safety risks that extend far beyond traditional LLM concerns, prompting a prolifera

The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure

SafetyDGX agent

arXiv:2605.17480v1 Announce Type: new Abstract: Multi-agent systems extend large language models (LLMs) by decomposing tasks among specialized agents, but their distributed decision process creates ne

Thinking with Patterns: Breaking the Perceptual Bottleneck in Visual Planning via Pattern Induction

Local AiDGX agent

arXiv:2605.16848v1 Announce Type: cross Abstract: Planning from raw visual input remains a significant challenge for current Vision-Language Models (VLMs), when the complexity of input is beyond their

TriALS: Triphasic-Aided Liver Lesion Segmentation Benchmark in Non-Contrast CT

Model ReleasesDGX agent

arXiv:2605.16572v1 Announce Type: new Abstract: Automated segmentation of liver lesions on non-contrast computed tomography (NCCT) is clinically important but fundamentally challenging, particularly i

Usenix'23 Extended Version: Smart Learning to Find Dumb Contracts

Model ReleasesDGX agent

arXiv:2304.10726v3 Announce Type: replace-cross Abstract: We introduce the Deep Learning Vulnerability Analyzer (DLVA) for Ethereum smart contracts based on neural networks. We train DLVA to judge byt

UVTran: Accurate Hole-Filling Parameterization with Transformers

Model ReleasesDGX agent

arXiv:2605.16306v1 Announce Type: cross Abstract: In industrial design, N-sided hole filling is typically formulated as the construction of a single trimmed B-spline surface by minimizing a fairness e

Validate Your Authority: Benchmarking LLMs on Multi-Label Precedent Treatment Classification

Model ReleasesDGX agent

arXiv:2605.17691v1 Announce Type: cross Abstract: Automating the classification of negative treatment in legal precedent is a critical yet nuanced NLP task where misclassification carries significant

VGGT-CD: Training-Free Robust Registration for 3D Change Detection

Model ReleasesDGX agent

arXiv:2605.16859v1 Announce Type: cross Abstract: 3D change detection from multi-view images is essential for urban monitoring, disaster assessment, and autonomous driving. However, existing methods p

Wavelet Flow Matching for Multi-Scale Physics Emulation

ResearchDGX agent

arXiv:2605.16573v1 Announce Type: cross Abstract: Accurate emulation of multi-scale physical systems governed by PDEs demands models that remain stable over long autoregressive rollouts while preservi

When Molecular Similarity Works: Property Cliffs Reveal Hidden Errors

Local AiDGX agent

arXiv:2605.17265v1 Announce Type: new Abstract: Accurate prediction of molecular properties underpins drug discovery and material design, yet even state-of-the-art models remain vulnerable to localize

Where Does Warm-Up Come From? Adaptive Scheduling for Norm-Constrained Optimizers

Model ReleasesDGX agent

arXiv:2602.05813v2 Announce Type: replace Abstract: We study adaptive learning rate scheduling for norm-constrained optimizers (e.g., Muon and Lion). We introduce a generalized smoothness assumption u

WhiteTesseract: Reframing the Interpretation of Cultural Heritage through XR and Conversational AI

Model ReleasesDGX agent

arXiv:2605.16972v1 Announce Type: cross Abstract: Cultural heritage exhibitions often struggle to sustain attention and support reflective engagement. Physical exhibitions rely on fixed interpretive a

WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens

Model ReleasesDGX agent

arXiv:2605.18115v1 Announce Type: new Abstract: Building a unified visual tokenizer is essential for bridging the gap between visual understanding and generation. Yet existing approaches struggle with

World understanding: Gemini Omni is built on Gemini's vast knowledge of history, science, and culture, so it can produce videos that are gro…

Model ReleasesDGX agent

Gemini Omni is built on Gemini's extensive knowledge base of history, science, and culture, enabling it to generate videos with sophisticated contextual understanding. The capability leverages Google'

18 May 2026

ActiveDPO: Active Direct Preference Optimization for Sample-Efficient Alignment

SafetyDGX agent

arXiv:2505.19241v2 Announce Type: replace-cross Abstract: The recent success in using human preferences to align large language models (LLMs) has significantly improved their performance in various do

AI-Mediated Communication Can Steer Collective Opinion

SafetyDGX agent

arXiv:2605.16245v1 Announce Type: cross Abstract: Generative artificial intelligence (AI) is increasingly integrated into the online platforms where humans exchange opinions; large language models (LL

ALSO: Adversarial Online Strategy Optimization for Social Agents

Model ReleasesDGX agent

arXiv:2605.15768v1 Announce Type: new Abstract: Social simulation provides a compelling testbed for studying social intelligence, where agents interact through multi-turn dialogues under evolving cont

Approximating Global Contact-Implicit MPC via Sampling and Local Complementarity

ResearchDGX agent

arXiv:2505.13350v2 Announce Type: replace Abstract: To achieve general-purpose dexterous manipulation, robots must rapidly devise and execute contact-rich behaviors. Existing model-based controllers a

Argus: Evidence Assembly for Scalable Deep Research Agents

Model ReleasesDGX agent

arXiv:2605.16217v1 Announce Type: cross Abstract: Deep research agents have achieved remarkable progress on complex information seeking tasks. Even long ReAct style rollouts explore only a single traj

Beyond First-Order: Learning Riemannian Geometries for Invariant Visual Place Recognition

Model ReleasesDGX agent

arXiv:2602.00841v4 Announce Type: replace Abstract: Visual Place Recognition (VPR) demands representations robust to drastic environmental and viewpoint shifts. Existing aggregation paradigms either d

Bidirectional Fusion Guided by Cardiac Patterns for Semi-Supervised ECG Segmentation

Model ReleasesDGX agent

arXiv:2605.15722v1 Announce Type: cross Abstract: Accurate delineation of electrocardiogram (ECG), the segmentation of meaningful waveform features, is crucial for cardiovascular diagnostics. However,

COCO-Inpaint: A Benchmark for Detecting and Localizing Inpainting-Based Image Manipulations

Model ReleasesDGX agent

arXiv:2504.18361v2 Announce Type: replace-cross Abstract: Recent advances in image manipulation have enabled highly photorealistic content generation, but also lowered the barrier to arbitrary editing

Context-aware Entity-Relation Extraction for Threat Intelligence Knowledge Graphs

Model ReleasesDGX agent

arXiv:2605.15904v1 Announce Type: new Abstract: Cybersecurity Knowledge Graphs (CKGs) unify diverse Cyber Threat Intelligence (CTI) sources into structured, queryable formats, offering scalable soluti

Context, Reasoning, and Hierarchy: A Cost-Performance Study of Compound LLM Agent Design in an Adversarial POMDP

AgentsDGX agent

arXiv:2605.16205v1 Announce Type: new Abstract: Deploying compound LLM agents in adversarial, partially observable sequential environments requires navigating several design dimensions: (1) what the a

DexJoCo: A Benchmark and Toolkit for Task-Oriented Dexterous Manipulation on MuJoCo

Model ReleasesDGX agent

arXiv:2605.16257v1 Announce Type: new Abstract: Achieving human-level manipulation requires dexterous robotic hands capable of complex object interactions. Advancing such capabilities further demands

DrugSAGE:Self-evolving Agent Experience for Efficient State-of-the-Art Drug Discovery

AgentsDGX agent

arXiv:2605.15461v1 Announce Type: cross Abstract: Building state-of-the-art (SOTA) predictive models for drug discovery requires expensive search over tools, architectures, and training strategies. Cu

FAR: Function-preserving Attention Replacement for IMC-friendly Inference

Local AiDGX agent

arXiv:2505.21535v4 Announce Type: replace-cross Abstract: While transformers dominate modern vision and language models, their attention mechanism remains poorly suited for in-memory computing (IMC) d

Flux Klein 9b in Easy diffusion

Local AiDGX agent

FLUX.2 [klein] 9B is a 9 billion parameter rectified flow transformer capable of generating images from text descriptions and supports multi-reference editing capabilities. The model matches or exceed

ForMaT: Dataset for Visually-Grounded Multilingual PDF Translation

Model ReleasesDGX agent

arXiv:2605.15794v1 Announce Type: new Abstract: We present ForMaT (Format-Preserving Multilingual Translation), a parallel corpus of 3,956 PDFs across 15 language pairs that preserves original layout

From Guidelines to Guarantees: A Graph-Based Evaluation Harness for Domain-Specific Evaluation of LLMs

ResearchDGX agent

arXiv:2508.20810v3 Announce Type: replace Abstract: Rigorous evaluation of domain-specific language models requires benchmarks that are comprehensive, contamination-resistant, and maintainable. Static

From Layers to Networks: Comparing Neural Representations via Diffusion Geometry

Model ReleasesDGX agent

arXiv:2605.15901v1 Announce Type: new Abstract: Diffusion geometry is a manifold learning framework that uses random walks defined by Markov transition matrices to characterize the geometry of a datas

H-Mem: A Novel Memory Mechanism for Evolving and Retrieving Agent Memory via a Hybrid Structure

AgentsDGX agent

arXiv:2605.15701v1 Announce Type: cross Abstract: Memory data are ubiquitous in Large Language Model (LLM)-based agents (e.g., OpenClaw and Manus). A few recent works have attempted to exploit agents'

Hardware-Software Co-Design of Scalable, Energy-Efficient Analog Recurrent Computations

Model ReleasesDGX agent

arXiv:2605.15216v1 Announce Type: cross Abstract: Always-on AI applications, from environmental sensors to biomedical implants, require ultra-low power consumption. Analog circuits offer a path to sub

HoloMotion-1 Technical Report

SafetyDGX agent

arXiv:2605.15336v1 Announce Type: cross Abstract: In this report, we present HoloMotion-1, a humanoid motion foundation model for zero-shot whole-body motion tracking. A key innovation of HoloMotion-1

Imitation learning for clinical decision support in pediatric ECMO

TutorialsDGX agent

arXiv:2605.16175v1 Announce Type: new Abstract: Pediatric critical care is a dynamic, high-stakes process involving constant monitoring and adjustments in life-saving treatments. Modeling these interv

Information-Preserving Domain Transfer with Unlabeled Data in Misspecified Simulation-Based Inference

Model ReleasesDGX agent

arXiv:2605.05652v2 Announce Type: replace Abstract: Simulation-based inference (SBI) provides amortized Bayesian parameter inference from simulator-generated data without requiring explicit likelihood

Interaction-Aware Influence Functions for Group Attribution

Model ReleasesDGX agent

arXiv:2605.15675v1 Announce Type: cross Abstract: Influence functions approximate how removing a training example changes a quantity of interest, called the target function, such as a held-out loss. T

Invaria: Learning Scale and Density Invariance in Point Clouds via Next-Resolution Prediction

TutorialsDGX agent

arXiv:2605.15923v1 Announce Type: new Abstract: Modern image encoders achieve high generalization by decoupling semantic meaning from resolution, an ability yet to be fully realized in the 3D domain.

ITGPT: Generative Pretraining on Irregular Timeseries

ApplicationsDGX agent

arXiv:2605.16069v1 Announce Type: new Abstract: Timeseries regression models often struggle to leverage large volumes of labeled multimodal data, particularly when the data are irregularly sampled or

JAM-Flow: Joint Audio-Motion Synthesis with Flow Matching

Local AiDGX agent

arXiv:2506.23552v2 Announce Type: replace Abstract: The intrinsic link between facial motion and speech is often overlooked in generative modeling, where talking head synthesis and text-to-speech (TTS

Layer-wise Derivative Controlled Networks

ApplicationsDGX agent

arXiv:2605.15463v1 Announce Type: new Abstract: As machine learning models grow in complexity, they increasingly struggle with three conflicting demands: the need for high accuracy, the requirement fo

LDGuid: A Framework for Robust Change Detection via Latent Difference Guidance

TutorialsDGX agent

arXiv:2605.15582v1 Announce Type: new Abstract: Modern deep learning models for change detection (CD) often struggle to explicitly represent task-relevant semantic differences. This paper proposes the

Learn2Splat: Extending the Horizon of Learned 3DGS Optimization

Model ReleasesDGX agent

arXiv:2605.15760v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) optimization is most commonly performed using standard optimizers (Adam, SGD). While stable across diverse scenes, standard

Learning Dynamic Structural Specialization for Underwater Salient Object Detection

Model ReleasesDGX agent

arXiv:2605.15535v1 Announce Type: new Abstract: Underwater salient object detection (USOD) has attracted increasing attention for underwater visual scene understanding and vision-guided robotic applic

Navigating Potholes with Geometry-Aware Sharpness Minimization

Model ReleasesDGX agent

arXiv:2605.16134v1 Announce Type: cross Abstract: Sharpness-aware minimization (SAM) encourages flat minima by perturbing parameters along directions of high loss curvature, but treats all parameter d

Position: Ideas Should be the Center of Machine Learning Research

Model ReleasesDGX agent

arXiv:2605.15253v1 Announce Type: new Abstract: Machine learning research increasingly bifurcates into two disconnected modes: benchmark-driven engineering that prioritizes metrics over understanding,

Prompting Amazon Nova 2 for content moderation

Model ReleasesDGX agent

In this post, you learn how to prompt Amazon Nova 2 Lite for content moderation using structured and free-form approaches, grounded in the MLCommons AILuminate Assessment Standard. The prompting techn

Registers Matter for Pixel-Space Diffusion Transformers

Model ReleasesDGX agent

arXiv:2605.16147v1 Announce Type: new Abstract: Vision Transformers (ViTs) are known to exhibit high-norm patch-token outliers that degrade feature map quality, a problem effectively mitigated by exti

Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems

SafetyDGX agent

arXiv:2605.15573v1 Announce Type: new Abstract: Multi-agent systems can solve complex tasks through collaboration between multiple Large Language Model agents. Existing collaboration frameworks typica

RoadmapBench: Evaluating Long-Horizon Agentic Software Development Across Version Upgrades

Model ReleasesDGX agent

arXiv:2605.15846v1 Announce Type: cross Abstract: Coding agents are increasingly deployed in real software development, where a single version iteration requires months of coordinated work across many

RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably

SafetyDGX agent

arXiv:2605.15514v1 Announce Type: cross Abstract: We identify intrinsic limitations of Rotary Positional Embeddings (RoPE) in Transformer-based long-context language models. Our theoretical analysis a

Run Claude Managed Agents with Vercel Sandbox

Model ReleasesDGX agent

This article describes how to run Claude's managed agents within Vercel's Sandbox environment, enabling developers to execute AI agent workloads on Vercel's infrastructure. The integration allows user

SDOF: Taming the Alignment Tax in Multi-Agent Orchestration with State-Constrained Dispatch

Model ReleasesDGX agent

arXiv:2605.15204v1 Announce Type: new Abstract: Multi-agent orchestration frameworks such as LangChain, LangGraph, and CrewAI route tasks through graph-based pipelines but do not enforce the stage con

Searching on a Budget: HW-NAS with 10 Latency Probes

Model ReleasesDGX agent

arXiv:2504.00663v2 Announce Type: replace Abstract: Existing hardware-aware NAS (HW-NAS) methods typically assume access to precise information circa the target device, either via analytical approxima

See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation

Local AiDGX agent

arXiv:2605.15585v1 Announce Type: new Abstract: Large language models can generate executable code for educational animations, but the resulting renders often exhibit visual defects, including element

Seeing is Understanding: Unlocking Causal Attention into Modality-Mutual Attention for Multimodal LLMs

ResearchDGX agent

arXiv:2503.02597v3 Announce Type: replace-cross Abstract: Recent Multimodal Large Language Models (MLLMs) have demonstrated significant progress in perceiving and reasoning over multimodal inquiries,

SOLAR: Self-supervised Joint Learning for Symmetric Multimodal Retrieval

Model ReleasesDGX agent

arXiv:2605.15868v1 Announce Type: new Abstract: In this work, we address the critical yet underexplored challenge of symmetric multimodal-to-multimodal (MM2MM) retrieval, where queries and contexts ar

StippleDiffusion: Capacity-Constrained Stippling using Controlled Diffusion

Model ReleasesDGX agent

arXiv:2605.15816v1 Announce Type: cross Abstract: Stipple patterns, point sets whose local density tracks a target image, are traditionally produced by per-density iterative optimizers, which are slow

← Previous
1…556557558559560…1061
Next →