AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

LMGenDrive: Bridging Multimodal Understanding and Generative World Modeling for End-to-End Driving

DGX agent

arXiv:2604.08719v1 Announce Type: cross Abstract: Recent years have seen remarkable progress in autonomous driving, yet generalization to long-tail and open-world scenarios remains a major bottleneck

safetyarxiv-cs-ai
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Mamba-Based Graph Convolutional Networks: Tackling Over-smoothing with Selective State Space

DGX agent

arXiv:2501.15461v4 Announce Type: replace Abstract: Graph Neural Networks (GNNs) have shown great success in various graph-based learning tasks. However, it often faces the issue of over-smoothing as

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability

DGX agent

arXiv:2604.08788v1 Announce Type: new Abstract: Patient-clinician communication is an asymmetric-information problem: patients often do not disclose fears, misconceptions, or practical barriers unless

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

Policy-Aware Design of Large-Scale Factorial Experiments

DGX agent

arXiv:2604.08804v1 Announce Type: cross Abstract: Digital firms routinely run many online experiments on shared user populations. When product decisions are compositional, such as combinations of inte

safetyarxiv-cs-lg
13 Apr 2026
Safety

RAMP: Hybrid DRL for Online Learning of Numeric Action Models

DGX agent

arXiv:2604.08685v1 Announce Type: new Abstract: Automated planning algorithms require an action model specifying the preconditions and effects of each action, but obtaining such a model is often hard.

safetyarxiv-cs-ai
13 Apr 2026
Research

Robust Adaptive Backstepping Impedance Control of Robots in Unknown Environments

DGX agent

arXiv:2604.09323v1 Announce Type: new Abstract: This paper presents a Robust Adaptive Backstepping Impedance Control (RABIC) strategy for robots operating in contact-rich and uncertain environments. T

researcharxiv-cs-ro
13 Apr 2026
Applications

SatQNet: Satellite-assisted Quantum Network Entanglement Routing Using Directed Line Graph Neural Networks

DGX agent

arXiv:2604.09306v1 Announce Type: cross Abstract: Quantum networks are expected to become a key enabler for interconnecting quantum devices. In contrast to classical communication networks, however, i

applicationsarxiv-cs-ai
13 Apr 2026
Safety

SPPO: Sequence-Level PPO for Long-Horizon Reasoning Tasks

DGX agent

arXiv:2604.08865v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) is central to aligning Large Language Models (LLMs) in reasoning tasks with verifiable rewards. However, standard tok

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

The AI Codebase Maturity Model: From Assisted Coding to Self-Sustaining Systems

DGX agent

arXiv:2604.09388v1 Announce Type: cross Abstract: AI coding tools are widely adopted, but most teams plateau at prompt-and-review without a framework for systematic progression. This paper presents th

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

DGX agent

arXiv:2601.23045v2 Announce Type: replace Abstract: As AI becomes more capable, we entrust it with more general and consequential tasks. The risks from failure grow more severe with increasing task sc

safetyarxiv-cs-ai
13 Apr 2026
Safety

The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMs

DGX agent

arXiv:2601.01580v2 Announce Type: replace-cross Abstract: Self-reflection capabilities emerge in Large Language Models after RL post-training, with multi-turn RL achieving substantial gains over SFT c

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

DGX agent

arXiv:2604.02241v2 Announce Type: replace Abstract: Embodied visual tracking is crucial for Unmanned Aerial Vehicles (UAVs) executing complex real-world tasks. In dynamic urban scenarios with complex

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails

DGX agent

arXiv:2603.03099v5 Announce Type: replace-cross Abstract: Despite Adam demonstrating faster empirical convergence than SGD in many applications, much of the existing theory yields guarantees essential

model-releasesarxiv-cs-ai
13 Apr 2026
Tutorials

WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning

DGX agent

arXiv:2604.08958v1 Announce Type: cross Abstract: Reinforcement learning (RL) in robotics is often limited by the cost and risk of data collection, motivating experience transfer from a source task to

tutorialsarxiv-cs-ai
13 Apr 2026
Model Releases

3DrawAgent: Teaching LLM to Draw in 3D with Early Contrastive Experience

DGX agent

arXiv:2604.08042v1 Announce Type: new Abstract: Sketching in 3D space enables expressive reasoning about shape, structure, and spatial relationships, yet generating 3D sketches through natural languag

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

A Giant-Step Baby-Step Classifier For Scalable and Real-Time Anomaly Detection In Industrial Control Systems and Water Treatment Systems

DGX agent

arXiv:2504.20906v4 Announce Type: replace-cross Abstract: The continuous monitoring of the interactions between cyber-physical components of any industrial control system (ICS) is required to secure a

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference

DGX agent

arXiv:2604.08133v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become a dominant architecture for scaling large language models due to their sparse activation mechanism. However, the s

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Asking like Socrates: Socrates helps VLMs understand remote sensing images

DGX agent

arXiv:2511.22396v2 Announce Type: replace-cross Abstract: Recent multimodal reasoning models, inspired by DeepSeek-R1, have significantly advanced vision-language systems. However, in remote sensing (

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

AutoReproduce: Automatic AI Experiment Reproduction with Paper Lineage

DGX agent

arXiv:2505.20662v3 Announce Type: replace Abstract: Efficient reproduction of research papers is pivotal to accelerating scientific progress. However, the increasing complexity of proposed methods oft

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Before We Trust Them: Decision-Making Failures in Navigation of Foundation Models

DGX agent

arXiv:2601.05529v5 Announce Type: replace Abstract: High success rates on navigation-related tasks do not necessarily translate into reliable decision making by foundation models. To examine this gap,

model-releasesarxiv-cs-ai
10 Apr 2026
Local Ai

Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding

DGX agent

arXiv:2604.08014v1 Announce Type: new Abstract: Spatio-Temporal Video Grounding requires jointly localizing target objects across both temporal and spatial dimensions based on natural language queries

local-aiarxiv-cs-cv
10 Apr 2026
Model Releases

Broken by Default: A Formal Verification Study of Security Vulnerabilities in AI-Generated Code

DGX agent

arXiv:2604.05292v2 Announce Type: replace-cross Abstract: AI coding assistants are now used to generate production code in security-sensitive domains, yet the exploitability of their outputs remains u

model-releasesarxiv-cs-ai
10 Apr 2026
Applications

CompoDistill: Attention Distillation for Compositional Reasoning in Multimodal LLMs

DGX agent

arXiv:2510.12184v2 Announce Type: replace Abstract: Recently, efficient Multimodal Large Language Models (MLLMs) have gained significant attention as a solution to their high computational complexity,

applicationsarxiv-cs-cv
10 Apr 2026
Model Releases

ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training

DGX agent

arXiv:2604.07484v1 Announce Type: cross Abstract: Generative reward models (GRMs) have emerged as a promising approach for aligning Large Language Models (LLMs) with human preferences by offering grea

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing

DGX agent

arXiv:2509.01986v4 Announce Type: replace-cross Abstract: In recent years, integrating multimodal understanding and generation into a single unified model has emerged as a promising paradigm. While th

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

DGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task

DGX agent

arXiv:2510.14509v3 Announce Type: replace-cross Abstract: The rapid advancement in large language models (LLMs) has demonstrated significant potential in End-to-End Software Development (E2ESD). Howev

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

DGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Evaluation as Evolution: Transforming Adversarial Diffusion into Closed-Loop Curricula for Autonomous Vehicles

DGX agent

arXiv:2604.07378v1 Announce Type: new Abstract: Autonomous vehicles in interactive traffic environments are often limited by the scarcity of safety-critical tail events in static datasets, which biase

safetyarxiv-cs-ro
10 Apr 2026
Model Releases

EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration

DGX agent

arXiv:2604.07070v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, their potential for purpose-driven exploration in dynamic geo-spatial

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Explainable AI to Improve Machine Learning Reliability for Industrial Cyber-Physical Systems

DGX agent

arXiv:2601.16074v2 Announce Type: replace Abstract: Industrial Cyber-Physical Systems (CPS) are sensitive infrastructure from both safety and economics perspectives, making their reliability criticall

safetyarxiv-cs-lg
10 Apr 2026
Applications

Exploring Natural Language-Based Strategies for Efficient Number Learning in Children through Reinforcement Learning

DGX agent

arXiv:2410.08334v2 Announce Type: replace-cross Abstract: In this paper, we build a reinforcement learning framework to study how children compose numbers using base-ten blocks. Studying numerical cog

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

FORGE:Fine-grained Multimodal Evaluation for Manufacturing Scenarios

DGX agent

arXiv:2604.07413v1 Announce Type: new Abstract: The manufacturing sector is increasingly adopting Multimodal Large Language Models (MLLMs) to transition from simple perception to autonomous execution,

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

From experimentation to engagement: on the paradox of participatory AI and power in contexts of forced displacement and humanitarian crises

DGX agent

arXiv:2604.06219v1 Announce Type: cross Abstract: Across the Global North, calls for participatory artificial intelligence (AI) to improve the responsible, safe, and ethical use of AI have increased,

safetyarxiv-cs-ai
10 Apr 2026
Safety

GIFT: Group-Relative Implicit Fine-Tuning Integrates GRPO with DPO and UNA

DGX agent

arXiv:2510.23868v4 Announce Type: replace Abstract: This paper proposes extit{Group-relative Implicit Fine-Tuning (GIFT)}, a reinforcement learning framework for aligning large language models (LLMs

safetyarxiv-cs-lg
10 Apr 2026
Hardware

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models

DGX agent

arXiv:2604.07812v1 Announce Type: new Abstract: In multimodal large language models (MLLMs), the surge of visual tokens significantly increases the inference time and computational overhead, making th

hardwarearxiv-cs-cv
10 Apr 2026
Applications

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models

DGX agent

arXiv:2512.09928v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently enabled robotic manipulation by grounding visual and linguistic cues into actions. However, most V

applicationsarxiv-cs-ro
10 Apr 2026
Research

Iteratively Learning Muscle Memory for Legged Robots to Master Adaptive and High Precision Locomotion

DGX agent

arXiv:2507.13662v2 Announce Type: replace Abstract: This paper presents a scalable and adaptive control framework for legged robots that integrates Iterative Learning Control (ILC) with a biologically

researcharxiv-cs-ro
10 Apr 2026
Research

LongSpec: Long-Context Lossless Speculative Decoding with Efficient Drafting and Verification

DGX agent

arXiv:2502.17421v4 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) can now process extremely long contexts, efficient inference over these extended inputs has become increasingl

researcharxiv-cs-ai
10 Apr 2026
Model Releases

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

DGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LPM 1.0: Video-based Character Performance Model

DGX agent

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference

DGX agent

arXiv:2509.22750v3 Announce Type: replace Abstract: Real-world multi-hop QA is naturally linked with ambiguity, where a single query can trigger multiple reasoning paths that require independent resol

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing

DGX agent

arXiv:2604.07747v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve low-k reasoning accuracy while narrowing solution coverage on challenging math que

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

MorphDistill: Distilling Unified Morphological Knowledge from Pathology Foundation Models for Colorectal Cancer Survival Prediction

DGX agent

arXiv:2604.06390v1 Announce Type: cross Abstract: Background: Colorectal cancer (CRC) remains a leading cause of cancer-related mortality worldwide. Accurate survival prediction is essential for treat

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Orion-Lite: Distilling LLM Reasoning into Efficient Vision-Only Driving Models

DGX agent

arXiv:2604.08266v1 Announce Type: new Abstract: Leveraging the general world knowledge of Large Language Models (LLMs) holds significant promise for improving the ability of autonomous driving systems

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Models

DGX agent

arXiv:2604.08340v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved remarkable progress in static visual understanding, their deployment in complex 3D embodied environmen

model-releasesarxiv-cs-cv
10 Apr 2026
Local Ai

Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models

DGX agent

arXiv:2604.06912v1 Announce Type: cross Abstract: MLLMs require high-resolution visual inputs for fine-grained tasks like document understanding and dense scene perception. However, current global res

local-aiarxiv-cs-ai
10 Apr 2026
Tutorials

Reading Recognition in the Wild

DGX agent

arXiv:2505.24848v4 Announce Type: replace Abstract: To enable egocentric contextual AI in always-on smart glasses, it is crucial to be able to keep a record of the user's interactions with the world,

tutorialsarxiv-cs-cv
10 Apr 2026
← Previous
1…230231232233
Next →