AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,259
  • Agents7,702
  • Applications5,506
  • Concepts5
  • Hardware1,896
  • Industry6,187
  • Local Ai5,048
  • Model Releases24,516
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,677
  • Tutorials3,454

Source
HumanDGX agent

Content type
90,259Total entries
1Added by human
90,258Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
64,475 results
Safety

Inverse Design of Realizable Metasurface based Absorbers using Improved Conditioning and Diversity Enhanced Progressively Growing GANs

DGX agent

arXiv:2606.05849v1 Announce Type: cross Abstract: Metasurfaces enable precise manipulation of electromagnetic waves for applications such as beam steering, sensing, and stealth technology. However, in

safetyarxiv-cs-cv
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Inverse Manipulation through Symbolic Planning and Residual Operator Learning

DGX agent

arXiv:2606.05248v1 Announce Type: new Abstract: Inverting a robotic task requires more than reversing symbolic state transitions or rewinding motor trajectories. In robot manipulation tasks, symbolic

safetyarxiv-cs-ro
5 Jun 2026
Research

IR3DE: A Linear Router for Large Language Models

DGX agent

arXiv:2606.06098v1 Announce Type: new Abstract: Foundational Large Language Models (LLMs) demonstrate proficiency on a wide range of general tasks, and achieve remarkable results on various specialize

researcharxiv-cs-cl
5 Jun 2026
Safety

Is Diversity All You Need for Scalable Robotic Manipulation?

DGX agent

arXiv:2507.06219v2 Announce Type: replace Abstract: Data scaling has driven remarkable success in foundation models for Natural Language Processing (NLP) and Computer Vision (CV), yet the principles o

safetyarxiv-cs-ro
5 Jun 2026
Model Releases

Is This Edit Correct? A Multi-Dimensional Benchmark for Reasoning-Aware Image Editing

DGX agent

arXiv:2606.05172v1 Announce Type: cross Abstract: Diffusion-based image editing has achieved strong visual fidelity under natural language instructions, yet most existing systems still operate at the

model-releasesarxiv-cs-cv
5 Jun 2026
Research

Knowledge Distillation for Visual Autoregressive Models

DGX agent

arXiv:2606.06078v1 Announce Type: new Abstract: Autoregressive (AR) image generation models are highly expressive but computationally intensive, motivating effective model compression. Knowledge disti

researcharxiv-cs-cv
5 Jun 2026
Model Releases

KV-Control: Parameter-Efficient K/V Injection for Trajectory-Controlled Text-to-Motion

DGX agent

arXiv:2606.05624v1 Announce Type: new Abstract: Text-conditioned 3D human motion models now synthesize plausible motions from prompts, but practical animation and embodied-agent workflows rarely stop

model-releasesarxiv-cs-cv
5 Jun 2026
Safety

L-SDPPO: Policy Optimization of Spiking Diffusion Policy for Intra-vehicular Robotic Manipulation

DGX agent

arXiv:2606.06049v1 Announce Type: new Abstract: Intra-vehicular robots in spacecraft help reduce astronaut workload and improve mission efficiency. Recent research focuses on using deep learning metho

safetyarxiv-cs-ro
5 Jun 2026
Safety

LadderMan: Learning Humanoid Perceptive Ladder Climbing

DGX agent

arXiv:2606.05873v1 Announce Type: cross Abstract: Humanoid robots hold great promise for operating in human-centered environments, yet ladder climbing remains one of the most challenging tasks due to

safetyarxiv-cs-cv
5 Jun 2026
Applications

LANTERN: Layered Archival and Temporal Episodic Retrieval Network for Long-Context LLM Conversations

DGX agent

arXiv:2606.05182v1 Announce Type: new Abstract: Large language models discard critical details when conversation history is compacted to fit within finite context windows. We present LANTERN (Layered

applicationsarxiv-cs-cl
5 Jun 2026
Safety

Large Language Models are Perplexed by some Political Parties

DGX agent

arXiv:2606.05937v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly used, including in political applications, but their political fairness has been little studied. We assess

safetyarxiv-cs-cl
5 Jun 2026
Research

Latent Implicit Visual Reasoning

DGX agent

arXiv:2512.21218v2 Announce Type: replace Abstract: While Large Multimodal Models (LMMs) have made significant progress, they remain largely text-centric, relying on language as their core reasoning m

researcharxiv-cs-cv
5 Jun 2026
Safety

Latent Reasoning with Normalizing Flows

DGX agent

arXiv:2606.06447v1 Announce Type: new Abstract: Large language models often improve reasoning by generating explicit chain-of-thought (CoT), demonstrating the importance of intermediate computation. H

safetyarxiv-cs-cl
5 Jun 2026
Model Releases

LatentSkill: From In-Context Textual Skills to In-Weight Latent Skills for LLM Agents

DGX agent

arXiv:2606.06087v1 Announce Type: new Abstract: Agent systems increasingly use textual skills to encode reusable task procedures, but injecting these skills into the prompt at every step incurs substa

model-releasesarxiv-cs-cl
5 Jun 2026
Local Ai

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization

DGX agent

arXiv:2606.05400v1 Announce Type: cross Abstract: Long-horizon autoformalization of research mathematics fails not only at hard lemmas, but at scale: statements drift, dependencies tangle, context dec

local-aiarxiv-cs-cl
5 Jun 2026
Research

Learning Contact Representation for Leg Odometry

DGX agent

arXiv:2606.05501v1 Announce Type: new Abstract: The estimation of odometry in legged robots depends on the assumption that the velocity of the foot with respect to the world remains zero during the st

researcharxiv-cs-ro
5 Jun 2026
Research

Learning from Demonstrations over Riemannian Manifolds using Neural ODEs: An Extended Abstract

DGX agent

arXiv:2606.05422v1 Announce Type: new Abstract: Learning from demonstratins (LfD) is usually performed over Euclidean spaces, while the robot state, e.g. orientation, naturally evolves over curved spa

researcharxiv-cs-ro
5 Jun 2026
Applications

Learning Geometric Representations from Videos for Spatial Intelligent Multimodal Large Language Models

DGX agent

arXiv:2606.05833v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) excel at 2D semantic understanding but lack intrinsic 3D awareness, resulting in representations that fail to m

applicationsarxiv-cs-cv
5 Jun 2026
Safety

Learning of Robot Safety Policies via Adversarial Synthetic Scenarios

DGX agent

arXiv:2606.05952v1 Announce Type: new Abstract: In this work, we propose an agentic gamification framework for hazard-informed learning of robot safety policies through synthetic scenarios. We model s

safetyarxiv-cs-ro
5 Jun 2026
Applications

Learning Predictive Visuomotor Coordination

DGX agent

arXiv:2503.23300v2 Announce Type: replace Abstract: Understanding and predicting human visuomotor coordination is crucial for applications in robotics, human-computer interaction, and assistive techno

applicationsarxiv-cs-cv
5 Jun 2026
Tutorials

Learning Self-Correction in Vision-Language Models via Rollout Augmentation

DGX agent

arXiv:2602.08503v2 Announce Type: replace-cross Abstract: Self-correction is essential for solving complex reasoning problems in vision-language models (VLMs). However, existing reinforcement learning

tutorialsarxiv-cs-cl
5 Jun 2026
Research

Learning to Route LLMs from Implicit Cost-Performance Preferences via Meta-Learning

DGX agent

arXiv:2606.06178v1 Announce Type: cross Abstract: Large language models (LLMs) present a trade-off between performance and cost, where more powerful models incur greater expense. LLM routing aims to m

researcharxiv-cs-cl
5 Jun 2026
Safety

Learning Visual Spatial Planning from Symbolic State via Modality-Gap-Aware Self-Distillation

DGX agent

arXiv:2606.06076v1 Announce Type: cross Abstract: While vision-language models excel at general multimodal understanding, they still struggle with visual spatial planning. We attribute this to a perce

safetyarxiv-cs-cv
5 Jun 2026
Research

Learning What to Forget: Improving LLM Unlearning via Learned Token-Level Importance

DGX agent

arXiv:2606.06320v1 Announce Type: cross Abstract: Machine unlearning aims to remove targeted knowledge from a trained model while preserving its general capabilities. For autoregressive language model

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Less is MoE: Trimming Experts in Domain-Specialist Language Models

DGX agent

arXiv:2606.05538v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) models achieve strong performance through conditional computation, but their large parameter footprint poses deployment chall

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

Let It Be Simple: One-Step Action Generation for Vision-Language-Action Models

DGX agent

arXiv:2606.05737v1 Announce Type: new Abstract: Diffusion-based vision-language-action (VLA) models often inherit the image-generation view: actions are generated by iterative denoising. We argue that

safetyarxiv-cs-cv
5 Jun 2026
Research

Leveraging Large Language Models for Generating Research Topic Ontologies: A Multi-Disciplinary Study

DGX agent

arXiv:2508.20693v2 Announce Type: replace-cross Abstract: Ontologies and taxonomies of research fields are critical for managing and organising scientific knowledge, as they facilitate efficient class

researcharxiv-cs-cl
5 Jun 2026
Model Releases

LiAuto-GeoX: Efficient Grounded Driving Transformer

DGX agent

arXiv:2606.05774v1 Announce Type: new Abstract: Dense 3D reconstruction has demonstrated immense potential for spatial understanding, yet its viability as a real-time, onboard representation for auton

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

LightVesselNet: An Ultra-Lightweight Sub-100K Parameter Network for Retinal Blood Vessel Segmentation

DGX agent

arXiv:2606.05354v1 Announce Type: new Abstract: Retinal blood vessel segmentation plays a vital role in the early detection of diabetic retinopathy and glaucoma. While recent deep learning models have

model-releasesarxiv-cs-cv
5 Jun 2026
Research

LLM-Conditioned Synthesis of Pathological Gaits via Structured Gait-Language Representations

DGX agent

arXiv:2606.06048v1 Announce Type: new Abstract: Pathological gait datasets remain scarce due to privacy, recruitment, cost, and movement variability. Our work presents a multimodal LLM-guided framewor

researcharxiv-cs-cv
5 Jun 2026
Research

LLM-Enhanced Dialogue Management for Full-Duplex Spoken Dialogue Systems

DGX agent

arXiv:2502.14145v3 Announce Type: replace Abstract: Achieving full-duplex communication in spoken dialogue systems (SDS) requires real-time coordination between listening, speaking, and thinking. This

researcharxiv-cs-cl
5 Jun 2026
Model Releases

LLM-Guided ANN Index Optimization for Human-Object Interaction Retrieval

DGX agent

arXiv:2606.05489v1 Announce Type: new Abstract: Retrieval systems underpin modern AI applications -- spanning visual search, recommendation engines, and multi-modal question answering. Modern multi-st

model-releasesarxiv-cs-cv
5 Jun 2026
Research

LLMs Can Leak Training Data But Do They Want To? A Propensity-Aware Evaluation of Memorization in LLMs

DGX agent

arXiv:2606.06286v1 Announce Type: new Abstract: Large language models can reproduce training data, but existing memorization evaluations mostly measure whether models can be forced to do so, rather th

researcharxiv-cs-cl
5 Jun 2026
Model Releases

Localizing Prompt Ambiguity in Large Language Models with Probe-Targeted Attribution

DGX agent

arXiv:2606.05486v1 Announce Type: new Abstract: Prompt ambiguity is a common source of failure in large language models, but is difficult to localize because it is a latent property of the prompt, whi

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

LongSpace: Exploring Long-Horizon Spatial Memory from Perception to Recall in Video

DGX agent

arXiv:2606.05677v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have advanced image and video understanding and can increasingly handle longer visual inputs. Long-horizon ta

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

LoomVideo: Unifying Multimodal Inputs into Video Generation and Editing

DGX agent

arXiv:2606.06042v1 Announce Type: new Abstract: Developing unified video generation and editing models capable of interpreting interleaved multimodal inputs is a promising yet challenging frontier fie

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

LoRi: Low-Rank Distillation for Implicit Reasoning

DGX agent

arXiv:2606.05315v1 Announce Type: new Abstract: Implicit chain-of-thought (iCoT) methods aim to internalize reasoning in large language models, but often underperform explicit CoT prompting. We empiri

model-releasesarxiv-cs-cl
5 Jun 2026
Research

Many Circuits, One Mechanism: Input Variation and Evaluation Granularity in Circuit Discovery

DGX agent

arXiv:2606.06267v1 Announce Type: new Abstract: Circuit discovery methods identify subgraphs that explain specific model behaviors, and structural differences between discovered circuits are commonly

researcharxiv-cs-cl
5 Jun 2026
Agents

MARDoc: A Memory-Aware Refinement Agent Framework for Multimodal Long Document QA

DGX agent

arXiv:2606.05749v1 Announce Type: new Abstract: Iterative retrieval-reasoning agents have recently shown promise for multimodal long-document question answering. However, most existing systems maintai

agentsarxiv-cs-cl
5 Jun 2026
Research

MASF: A Multi-Model Adaptive Selection Framework for Abstractive Text summarization

DGX agent

arXiv:2606.05494v1 Announce Type: new Abstract: Automatic text summarization has become increasingly important due to the rapid growth of digital textual information. This paper presents a Multi-Model

researcharxiv-cs-cl
5 Jun 2026
Model Releases

MAviS: A Multimodal Conversational Assistant For Avian Species

DGX agent

arXiv:2603.07294v2 Announce Type: replace Abstract: Fine-grained understanding and species-specific multimodal question answering are vital for advancing biodiversity conservation and ecological monit

model-releasesarxiv-cs-cv
5 Jun 2026
Model Releases

MCBench: A Multicontext Safety Assessment Benchmark for Omni Large Language Models

DGX agent

arXiv:2606.05177v1 Announce Type: new Abstract: Existing multimodal safety benchmarks focus solely on visual inputs and cannot assess Omni Large Language Models (LLMs) that process vision, audio, and

model-releasesarxiv-cs-cl
5 Jun 2026
Model Releases

MDP-GRPO: Stabilized Group Relative Policy Optimization for Multi-Constraint Instruction Following

DGX agent

arXiv:2606.06058v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards is ideal for multi-constraint instruction following, yet standard group-relative policy optimization (G

model-releasesarxiv-cs-cl
5 Jun 2026
Research

Measuring the sensitivity of LLM-based structured extraction to prompt, model, and schema choices in clinical discharge summaries

DGX agent

arXiv:2606.05970v1 Announce Type: new Abstract: Large language models are increasingly used for structured extraction from clinical free-text notes, but the sensitivity of their output to upstream con

researcharxiv-cs-cl
5 Jun 2026
Local Ai

Mechanistic Insights into Functional Sparsity in Multimodal LLMs via CoRe Heads

DGX agent

arXiv:2606.05843v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) demonstrate remarkable proficiency on complex vision-language tasks, the mechanisms by which they extract

local-aiarxiv-cs-cl
5 Jun 2026
Safety

Membrane: A Self-Evolving Contrastive Safety Memory for LLM Agent Defense

DGX agent

arXiv:2606.05743v1 Announce Type: cross Abstract: Despite advances in safety alignment, large language models remain vulnerable to continuously evolving jailbreaks. Existing fine-tuned safety classifi

safetyarxiv-cs-cl
5 Jun 2026
Research

MemoryCard: Topic-Aware Multi-Modal Clue Compression for Long-Video Question Answering

DGX agent

arXiv:2606.05917v1 Announce Type: cross Abstract: Long-video question answering remains challenging for Vision-Language Models (VLMs), as answer-relevant evidence is often sparse, transient, and tempo

researcharxiv-cs-cl
5 Jun 2026
Agents

Merging model-based control with multi-agent reinforcement learning for multi-agent cooperative teaming strategies

DGX agent

arXiv:2606.06011v1 Announce Type: new Abstract: In this work, we propose a framework that combines multi-agent reinforcement learning (MARL) with model-based control to achieve safe, dynamically feasi

agentsarxiv-cs-ro
5 Jun 2026
← Previous
1…626627628629630…1344
Next →