AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,356
  • Agents7,554
  • Applications5,409
  • Concepts5
  • Hardware1,835
  • Industry6,165
  • Local Ai4,929
  • Model Releases23,869
  • Research20,124
  • Safety13,369
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

88,356Total entries
1Added by human
88,355Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,584 results
2 Jun 2026

APE: Agentic Prompt Enhancer for Image Generation and Editing

Model ReleasesDGX agent

arXiv:2606.00204v1 Announce Type: new Abstract: Natural language has become a powerful interface for image generation and editing, yet text-guided visual systems remain highly sensitive to prompt form

ASE-26: a curriculum for agentic software engineering as a discipline

Model ReleasesDGX agent

arXiv:2606.01152v1 Announce Type: cross Abstract: The work of a professional software engineer has begun to consist, increasingly, of directing agents rather than writing code, and the empirical evide

Beyond Static Gaussians: An Empirical Investigation of Architectural Paradigms for Dynamic 3D Scene Reconstruction

Model ReleasesDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.00452v1 Announce Type: new Abstract: Dynamic scene reconstruction via 3D Gaussian Splatting (3DGS) has emerged as a compelling approach for representing evolving environments, yet understan

Bridging Reasoning Trajectories in On-Policy Distillation via Near-Future Guidance

SafetyDGX agent

arXiv:2606.00305v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) improves large language model reasoning by training a student model on trajectories sampled from its own policy under tea

Bridging Requirements and Architecture: Multi-Agent Orchestration with External Knowledge and Hierarchical Memory

Model ReleasesDGX agent

arXiv:2606.01385v1 Announce Type: cross Abstract: Software architecture design is a critical yet inherently complex and knowledge-intensive phase that requires balancing competing quality attributes a

CardioLens: Revealing the Clinical Reality Gap of MLLMs via Multi-Sequence Cardiac MRI Evaluations

ApplicationsDGX agent

arXiv:2606.00123v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have shown strong performance on public medical benchmarks, yet existing evaluations often remain weak proxie

ChatUMM: Robust Context Tracking for Conversational Interleaved Generation

ResearchDGX agent

arXiv:2602.06442v2 Announce Type: replace Abstract: Unified multimodal models (UMMs) have achieved remarkable progress yet remain constrained by a single-turn interaction paradigm, effectively functio

ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

Model ReleasesDGX agent

arXiv:2606.01494v1 Announce Type: cross Abstract: Agent skills extend AI agents with reusable instructions, tools, scripts, references, and workflows, establishing a security boundary distinct from bo

Coarse-to-Fine Compositional Diffusion for Long-Horizon Planning

Local AiDGX agent

arXiv:2606.00837v1 Announce Type: cross Abstract: Diffusion models provide strong priors for generating structured data, but many tasks require outputs beyond the scale on which these models are typic

CRePE: Convolution-aware Relative Importance in Post-training Pruning with Efficient Search

Local AiDGX agent

arXiv:2606.01544v1 Announce Type: new Abstract: Deploying Large Language Models (LLMs) in practice incurs substantial memory and computational costs. Post-training pruning (PTP) is an effective approa

Critic-R: Improving Agentic Search using Instruction-tuned Retrievers with Natural Language Introspective Feedback

AgentsDGX agent

arXiv:2606.00590v1 Announce Type: cross Abstract: Agentic search systems iteratively interact with retrieval models to answer complex queries. Despite substantial progress, optimizing retrievers for a

DAG-Plan: Generating Directed Acyclic Dependency Graphs for Dual-Arm Cooperative Planning

Model ReleasesDGX agent

arXiv:2406.09953v4 Announce Type: replace-cross Abstract: Dual-arm robots promise greater efficiency but require planning for complex tasks with nonlinear sub-task dependencies. Current methods using

DAStatFormer: A Hybrid Multibranch Transformer with Statistical Feature Integration for DAS-Based Pattern Recognitions

Model ReleasesDGX agent

arXiv:2606.00081v1 Announce Type: cross Abstract: Distributed Acoustic Sensing (DAS) enables large-scale monitoring through optical fibers, but its high dimensionality and complex spatio-temporal patt

DECK: A Consistency x Confidence Taxonomy of LLM Hallucinations

Model ReleasesDGX agent

arXiv:2606.02289v1 Announce Type: new Abstract: Existing hallucination taxonomies classify LLM errors by what is wrong with the output -- memorised misconceptions, reasoning failures, fluent fabricati

Deep Research as Rubric for Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.01091v1 Announce Type: new Abstract: Open-ended reasoning and long-form generation tasks lack reliable automatic verification signals for reward-based policy optimization. Rubrics offer a p

DFlare: Scaling Up Draft Capacity for Block Diffusion Speculative Decoding

ResearchDGX agent

arXiv:2606.02091v1 Announce Type: new Abstract: Block diffusion speculative decoding accelerates LLM inference by predicting all tokens within a block simultaneously for the target model to verify in

Diagnosing LLM Arbitration Behavior over Pre-evidence Epistemic States in RAG-based Fact-Checking

Model ReleasesDGX agent

arXiv:2606.01120v1 Announce Type: new Abstract: In RAG-based fact-checking, LLMs are increasingly used as verifiers to check given claims against retrieved evidence. Their parametric knowledge can ind

DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation

ResearchDGX agent

arXiv:2606.00535v1 Announce Type: new Abstract: Speculative decoding (SD) has proven to be an effective technique for accelerating autoregressive generation in large language models (LLMs) however, it

DSL-LLaDA: Scaling Continuous Denoising to 8B Masked Diffusion LMs

Local AiDGX agent

arXiv:2606.01024v1 Announce Type: cross Abstract: Discrete Masked diffusion language models generate text by iterative parallel decoding, but few-step decoding suffers from a tradeoff between length a

Early Diagnosis of Wasted Computation in Multi-Agent LLM Systems via Failure-Aware Observability

AgentsDGX agent

arXiv:2606.01365v1 Announce Type: new Abstract: Tool-using multi-agent large language model (LLM) systems spend computation through model tokens, tool calls, retries, and code execution before produci

Easier to Mislead Than to Correct: Harmful and Beneficial Revision in LLM Conformity

AgentsDGX agent

arXiv:2606.01637v1 Announce Type: cross Abstract: Large language models are increasingly used in multi-agent systems, where they see and respond to other agents' answers. A key risk is conformity: a m

EMoE: Training-Free Expert Disagreement for Uncertainty-Aware Text-to-Image Diffusion

SafetyDGX agent

arXiv:2505.13273v2 Announce Type: replace Abstract: Large text-to-image diffusion models rarely expose reliable signals of when a prompt is likely to produce a poorly aligned generation, especially wh

Evaluating and Learning Robust Bandit Policies Under Uncertain Causal Mechanisms

SafetyDGX agent

arXiv:2508.02812v3 Announce Type: replace Abstract: Causal graphical models can encode large amounts structural knowledge, both from the background knowledge of domain experts and the structural knowl

EvoPool: Evolutionary Programmatic Annotation for Label-Efficient Specialized Supervision

AgentsDGX agent

arXiv:2606.01617v1 Announce Type: cross Abstract: Large language models excel at general tasks but underperform smaller supervised models in specialized, high-stakes domains where training labels are

FigSIM: A Dataset for Fine-grained Suicide Severity and Figurative Language in Suicide Memes

Model ReleasesDGX agent

arXiv:2606.02523v1 Announce Type: new Abstract: Suicide memes are memes used to express suicide-related thoughts or comment on suicide-related issues. Suicide memes are increasingly common on social m

Flowers: A Warp Drive for Neural PDE Solvers

Model ReleasesDGX agent

arXiv:2603.04430v2 Announce Type: replace Abstract: We introduce Flowers, a neural architecture for learning PDE solution operators built entirely from multihead warps. Aside from pointwise channel mi

FlowIt: Global Matching via Hierarchical Transformers and Optimal Transport for Optical Flow

Model ReleasesDGX agent

arXiv:2603.28759v2 Announce Type: replace Abstract: We present FlowIt, a novel architecture for optical flow estimation that combines global matching with confidence and occlusion-guided refinement. A

FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection

Model ReleasesDGX agent

arXiv:2606.00782v1 Announce Type: new Abstract: Open-vocabulary object detection (OVD) has achieved remarkable progress through large-scale vision-language pre-training. Existing methods, however, typ

Forget Attention: Importance-Aware Attention Is All You Need

ResearchDGX agent

arXiv:2606.02332v1 Announce Type: new Abstract: Combining attention's global retrieval with the sequential importance signal of state space models (SSMs) is the open challenge of hybrid language model

From Empathy to Personalized Empathy: Adapting Empathetic Strategies to Individual Users

ResearchDGX agent

arXiv:2606.00728v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in long-term interactions with users, empathy has become an increasingly important capability.

From Human Videos to Robot Manipulation: A Survey on Scalable Vision-Language-Action Learning with Human-Centric Data

ApplicationsDGX agent

arXiv:2606.00054v1 Announce Type: cross Abstract: Recent progress in generalizable embodied control has been driven by large-scale pretraining of Vision-Language-Action (VLA) models. However, most exi

From Noise to Order: Learning to Rank via Denoising Diffusion

ResearchDGX agent

arXiv:2602.11453v2 Announce Type: replace-cross Abstract: In information retrieval (IR), learning-to-rank (LTR) methods have traditionally limited themselves to discriminative machine learning approac

From Rashomon Theory to PRAXIS: Efficient Decision Tree Rashomon Sets

ApplicationsDGX agent

arXiv:2606.00202v1 Announce Type: cross Abstract: Standard machine learning pipelines often admit many near-optimal models. These 'Rashomon sets' pose a range of challenges and opportunities for uncer

G2LoRA: Gradient Orthogonal Low-Rank Adaptation Framework for Graph Continual Learning on Text-Attributed Graphs

Model ReleasesDGX agent

arXiv:2606.01873v1 Announce Type: new Abstract: LLM-as-Aligner has emerged as a prevalent pre-training paradigm for Text-Attributed Graphs(TAGS), aligning graph and text modalities into a shared embed

Geodesics with Unified Tangent-constrained Priors and Curvature Regularization

ResearchDGX agent

arXiv:2606.00139v1 Announce Type: cross Abstract: Curvature-penalized geodesic models have proven their effectiveness in image segmentation by computing globally optimal curves. Unfortunately, these m

Geometric Latent Reasoning Induces Shorter Generations in LLMs

ResearchDGX agent

arXiv:2606.02248v1 Announce Type: new Abstract: Large language models solve complex problems by generating lengthy chains of explicit reasoning tokens. While effective, this makes reasoning expensive,

GLIDE: Graph-guided Leap Inference for Diffusion Estimation of Spatio-Temporal Point Processes

Local AiDGX agent

arXiv:2606.01273v1 Announce Type: new Abstract: Spatio-temporal point processes (STPPs) provide a principled framework for modeling asynchronous events in continuous time and space. Recent diffusion-b

Graph Edit Distance Formulation for the Vehicle Routing Problem: Theory and Analysis

Model ReleasesDGX agent

arXiv:2606.01987v1 Announce Type: cross Abstract: We show that the Vehicle Routing Problem (VRP) can be reformulated as a Graph Edit Distance (GED) maximization problem. Under a simple edge-deletion c

GuidaPA: Privacy-Preserving Chatbot for Public Administration via Federated Learning

Model ReleasesDGX agent

arXiv:2606.01386v1 Announce Type: new Abstract: We present GuidaPA, a privacy-preserving chatbot for the Italian Public Administration (PA) trained via Federated Learning (FL) on documentation from tw

HERO'S JOURNEY: Testing Complex Rule Induction with Text Games

Model ReleasesDGX agent

arXiv:2606.02556v1 Announce Type: new Abstract: We introduce HERO'S JOURNEY, a benchmark for rule induction in goal-directed episodic tasks, where agents must infer hidden rules from demonstrations an

Hierarchically Decoupled Mixture-of-Experts for Robust Traffic Sign Recognition in Complex Driving Scenarios

Model ReleasesDGX agent

arXiv:2606.01822v1 Announce Type: new Abstract: Traffic sign detection is a fundamental component of environmental perception in autonomous driving and intelligent transportation systems. However, mos

Human in the Loop Adaptive Optimization for Improved Time Series Forecasting

TutorialsDGX agent

arXiv:2505.15354v2 Announce Type: replace Abstract: Time series forecasting models often produce systematic, predictable errors even in critical domains such as energy, finance, and healthcare. We int

'I Strongly Suspect This Website Is a Scam': Benchmarking PII Leakage and Detection without Defense in Autonomous Web Agents

Model ReleasesDGX agent

arXiv:2606.00497v1 Announce Type: cross Abstract: Deceptive web content, widely instantiated across the internet and commonly known as extit{social-engineering attacks}, manipulates autonomous web age

IDEAFix: Evaluation Framework for Creative Defixation Prompting in LLMs

ResearchDGX agent

arXiv:2606.00875v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for tasks involving creative problem solving and idea generation. However, there is a lack of consens

Internalize the Temperature: On-Policy Self-Distillation as Policy Reheater for Reinforcement Learning

SafetyDGX agent

arXiv:2606.00755v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards improves the reasoning ability of large language models, but often suffers from entropy collapse, in whic

Inverse Depth Scaling From Most Layers Being Similar

SafetyDGX agent

arXiv:2602.05970v2 Announce Type: replace-cross Abstract: Neural scaling laws relate loss to model size in large language models (LLMs), yet depth and width may contribute to performance differently,

JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions

Model ReleasesDGX agent

arXiv:2606.01703v1 Announce Type: cross Abstract: We address the challenge of generating high-fidelity, long-form soundtracks that remain coherent across scene transitions. Existing AI music systems a

Knowledge-Intensive Video Generation

Model ReleasesDGX agent

arXiv:2606.01285v1 Announce Type: cross Abstract: Text-to-video generation has advanced rapidly in visual quality, but remains under-evaluated for factuality and practical usefulness. We introduce kno

Latent Diffusion Pretraining for Crystal Property Prediction

Local AiDGX agent

arXiv:2606.00776v1 Announce Type: new Abstract: Fast and accurate prediction of crystal properties is a central challenge in new materials design. Graph neural networks and Transformer-based models ha

Learning-based Directed Graph Abstraction of Combinatorial Spaces for Order-Preserving Search in Mixed-Combinatorial Nonlinear Optimization

Model ReleasesDGX agent

arXiv:2606.01425v1 Announce Type: new Abstract: Mixed-combinatorial nonlinear programming (MCNLP) problems arise in many engineering design and planning applications, e.g., due to categorical, compone

Local Diagnostics of Continuous Normalizing Flow for Out-of-Distribution Detection

Model ReleasesDGX agent

arXiv:2606.00684v1 Announce Type: cross Abstract: We address the problem of out-of-distribution (OOD) detection for target observations embedded in a subspace of the high dimensional data space. Using

Masking Stale Observations Helps Search Agents -- Until It Doesn't: A Regime Map and Its Mechanism

AgentsDGX agent

arXiv:2606.00408v1 Announce Type: cross Abstract: Long-horizon search agents accumulate large amounts of retrieved content across many tool calls, making context-budget efficiency increasingly importa

Measurement-Driven Early Warning of Reliability Breakdown in 5G NSA Railway Networks

Model ReleasesDGX agent

arXiv:2511.08851v5 Announce Type: replace-cross Abstract: This paper presents a measurement-driven study of early warning for reliability breakdown events in 5G non-standalone (NSA) railway networks.

MedGym:A Unified Continuous-Time Benchmark for Dynamic Medical Treatment Reinforcement Learning

Model ReleasesDGX agent

arXiv:2606.01028v1 Announce Type: new Abstract: Medical treatment recommendation poses several challenges to reinforcement learning (RL): patient physiology evolves in continuous time, measurements an

MindZero: Learning Online Mental Reasoning With Zero Annotations

ApplicationsDGX agent

arXiv:2606.00240v1 Announce Type: new Abstract: Effective real-world assistance requires AI agents with robust Theory of Mind (ToM): inferring human mental states from their behavior. Despite recent a

MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?

Model ReleasesDGX agent

arXiv:2606.01993v1 Announce Type: cross Abstract: Abundant procedural knowledge on the Web holds great potential for helping agents solve long-horizon tasks. However, such knowledge is often multimoda

MotionDreamer: Universal Skeletal Motion Generation for 3D Rigged Shapes

Model ReleasesDGX agent

arXiv:2606.01518v1 Announce Type: new Abstract: Motion generation for rigged shapes is vital for scalable 4D asset production. However, template-based methods are limited by specific topologies and fa

Multimodal Music Recommendation System using LLMs

Model ReleasesDGX agent

arXiv:2606.00125v1 Announce Type: cross Abstract: Music recommendation systems typically treat songs as opaque tokens, relying on collaborative interaction histories which overlooks semantic or acoust

Neural Network Compression by Approximate Differential Equivalence

Model ReleasesDGX agent

arXiv:2606.01402v1 Announce Type: cross Abstract: Neural network compression is commonly achieved by pruning parameters based on local importance scores, e.g., magnitude-based pruning. We propose a co

NVIDIA Partners With Microsoft on Unified Stack for Agentic AI Deployment, From Windows Devices to Cloud to Local

Local AiDGX agent

The agentic AI moment has arrived, but delivering on its promise requires more than good models. It also takes fast hardware, secure runtimes, a responsive data layer and models tuned for long-running

← Previous
1…464465466467468…1060
Next →