AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

The AI Codebase Maturity Model: From Assisted Coding to Self-Sustaining Systems

DGX agent

arXiv:2604.09388v1 Announce Type: cross Abstract: AI coding tools are widely adopted, but most teams plateau at prompt-and-review without a framework for systematic progression. This paper presents th

model-releasesarxiv-cs-ai
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

The Hot Mess of AI: How Does Misalignment Scale With Model Intelligence and Task Complexity?

DGX agent

arXiv:2601.23045v2 Announce Type: replace Abstract: As AI becomes more capable, we entrust it with more general and consequential tasks. The risks from failure grow more severe with increasing task sc

safetyarxiv-cs-ai
13 Apr 2026
Safety

The Two-Stage Decision-Sampling Hypothesis: Understanding the Emergence of Self-Reflection in RL-Trained LLMs

DGX agent

arXiv:2601.01580v2 Announce Type: replace-cross Abstract: Self-reflection capabilities emerge in Large Language Models after RL post-training, with multi-turn RL achieving substantial gains over SFT c

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

UAV-Track VLA: Embodied Aerial Tracking via Vision-Language-Action Models

DGX agent

arXiv:2604.02241v2 Announce Type: replace Abstract: Embodied visual tracking is crucial for Unmanned Aerial Vehicles (UAVs) executing complex real-world tasks. In dynamic urban scenarios with complex

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Why Adam Can Beat SGD: Second-Moment Normalization Yields Sharper Tails

DGX agent

arXiv:2603.03099v5 Announce Type: replace-cross Abstract: Despite Adam demonstrating faster empirical convergence than SGD in many applications, much of the existing theory yields guarantees essential

model-releasesarxiv-cs-ai
13 Apr 2026
Tutorials

WOMBET: World Model-based Experience Transfer for Robust and Sample-efficient Reinforcement Learning

DGX agent

arXiv:2604.08958v1 Announce Type: cross Abstract: Reinforcement learning (RL) in robotics is often limited by the cost and risk of data collection, motivating experience transfer from a source task to

tutorialsarxiv-cs-ai
13 Apr 2026
Model Releases

3DrawAgent: Teaching LLM to Draw in 3D with Early Contrastive Experience

DGX agent

arXiv:2604.08042v1 Announce Type: new Abstract: Sketching in 3D space enables expressive reasoning about shape, structure, and spatial relationships, yet generating 3D sketches through natural languag

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

A Giant-Step Baby-Step Classifier For Scalable and Real-Time Anomaly Detection In Industrial Control Systems and Water Treatment Systems

DGX agent

arXiv:2504.20906v4 Announce Type: replace-cross Abstract: The continuous monitoring of the interactions between cyber-physical components of any industrial control system (ICS) is required to secure a

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

Alloc-MoE: Budget-Aware Expert Activation Allocation for Efficient Mixture-of-Experts Inference

DGX agent

arXiv:2604.08133v1 Announce Type: cross Abstract: Mixture-of-Experts (MoE) has become a dominant architecture for scaling large language models due to their sparse activation mechanism. However, the s

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Asking like Socrates: Socrates helps VLMs understand remote sensing images

DGX agent

arXiv:2511.22396v2 Announce Type: replace-cross Abstract: Recent multimodal reasoning models, inspired by DeepSeek-R1, have significantly advanced vision-language systems. However, in remote sensing (

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

AutoReproduce: Automatic AI Experiment Reproduction with Paper Lineage

DGX agent

arXiv:2505.20662v3 Announce Type: replace Abstract: Efficient reproduction of research papers is pivotal to accelerating scientific progress. However, the increasing complexity of proposed methods oft

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Before We Trust Them: Decision-Making Failures in Navigation of Foundation Models

DGX agent

arXiv:2601.05529v5 Announce Type: replace Abstract: High success rates on navigation-related tasks do not necessarily translate into reliable decision making by foundation models. To examine this gap,

model-releasesarxiv-cs-ai
10 Apr 2026
Local Ai

Bridging Time and Space: Decoupled Spatio-Temporal Alignment for Video Grounding

DGX agent

arXiv:2604.08014v1 Announce Type: new Abstract: Spatio-Temporal Video Grounding requires jointly localizing target objects across both temporal and spatial dimensions based on natural language queries

local-aiarxiv-cs-cv
10 Apr 2026
Model Releases

Broken by Default: A Formal Verification Study of Security Vulnerabilities in AI-Generated Code

DGX agent

arXiv:2604.05292v2 Announce Type: replace-cross Abstract: AI coding assistants are now used to generate production code in security-sensitive domains, yet the exploitability of their outputs remains u

model-releasesarxiv-cs-ai
10 Apr 2026
Applications

CompoDistill: Attention Distillation for Compositional Reasoning in Multimodal LLMs

DGX agent

arXiv:2510.12184v2 Announce Type: replace Abstract: Recently, efficient Multimodal Large Language Models (MLLMs) have gained significant attention as a solution to their high computational complexity,

applicationsarxiv-cs-cv
10 Apr 2026
Model Releases

ConsistRM: Improving Generative Reward Models via Consistency-Aware Self-Training

DGX agent

arXiv:2604.07484v1 Announce Type: cross Abstract: Generative reward models (GRMs) have emerged as a promising approach for aligning Large Language Models (LLMs) with human preferences by offering grea

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Draw-In-Mind: Rebalancing Designer-Painter Roles in Unified Multimodal Models Benefits Image Editing

DGX agent

arXiv:2509.01986v4 Announce Type: replace-cross Abstract: In recent years, integrating multimodal understanding and generation into a single unified model has emerged as a promising paradigm. While th

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

DSCA: Dynamic Subspace Concept Alignment for Lifelong VLM Editing

DGX agent

arXiv:2604.07965v1 Announce Type: new Abstract: Model editing aims to update knowledge to add new concepts and change relevant information without retraining. Lifelong editing is a challenging task, p

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task

DGX agent

arXiv:2510.14509v3 Announce Type: replace-cross Abstract: The rapid advancement in large language models (LLMs) has demonstrated significant potential in End-to-End Software Development (E2ESD). Howev

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Efficient and Effective Internal Memory Retrieval for LLM-Based Healthcare Prediction

DGX agent

arXiv:2604.07659v1 Announce Type: new Abstract: Large language models (LLMs) hold significant promise for healthcare, yet their reliability in high-stakes clinical settings is often compromised by hal

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Evaluation as Evolution: Transforming Adversarial Diffusion into Closed-Loop Curricula for Autonomous Vehicles

DGX agent

arXiv:2604.07378v1 Announce Type: new Abstract: Autonomous vehicles in interactive traffic environments are often limited by the scarcity of safety-critical tail events in static datasets, which biase

safetyarxiv-cs-ro
10 Apr 2026
Model Releases

EVGeoQA: Benchmarking LLMs on Dynamic, Multi-Objective Geo-Spatial Exploration

DGX agent

arXiv:2604.07070v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, their potential for purpose-driven exploration in dynamic geo-spatial

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Explainable AI to Improve Machine Learning Reliability for Industrial Cyber-Physical Systems

DGX agent

arXiv:2601.16074v2 Announce Type: replace Abstract: Industrial Cyber-Physical Systems (CPS) are sensitive infrastructure from both safety and economics perspectives, making their reliability criticall

safetyarxiv-cs-lg
10 Apr 2026
Applications

Exploring Natural Language-Based Strategies for Efficient Number Learning in Children through Reinforcement Learning

DGX agent

arXiv:2410.08334v2 Announce Type: replace-cross Abstract: In this paper, we build a reinforcement learning framework to study how children compose numbers using base-ten blocks. Studying numerical cog

applicationsarxiv-cs-ai
10 Apr 2026
Model Releases

FORGE:Fine-grained Multimodal Evaluation for Manufacturing Scenarios

DGX agent

arXiv:2604.07413v1 Announce Type: new Abstract: The manufacturing sector is increasingly adopting Multimodal Large Language Models (MLLMs) to transition from simple perception to autonomous execution,

model-releasesarxiv-cs-cv
10 Apr 2026
Safety

From experimentation to engagement: on the paradox of participatory AI and power in contexts of forced displacement and humanitarian crises

DGX agent

arXiv:2604.06219v1 Announce Type: cross Abstract: Across the Global North, calls for participatory artificial intelligence (AI) to improve the responsible, safe, and ethical use of AI have increased,

safetyarxiv-cs-ai
10 Apr 2026
Safety

GIFT: Group-Relative Implicit Fine-Tuning Integrates GRPO with DPO and UNA

DGX agent

arXiv:2510.23868v4 Announce Type: replace Abstract: This paper proposes extit{Group-relative Implicit Fine-Tuning (GIFT)}, a reinforcement learning framework for aligning large language models (LLMs

safetyarxiv-cs-lg
10 Apr 2026
Hardware

HAWK: Head Importance-Aware Visual Token Pruning in Multimodal Models

DGX agent

arXiv:2604.07812v1 Announce Type: new Abstract: In multimodal large language models (MLLMs), the surge of visual tokens significantly increases the inference time and computational overhead, making th

hardwarearxiv-cs-cv
10 Apr 2026
Applications

HiF-VLA: Hindsight, Insight and Foresight through Motion Representation for Vision-Language-Action Models

DGX agent

arXiv:2512.09928v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently enabled robotic manipulation by grounding visual and linguistic cues into actions. However, most V

applicationsarxiv-cs-ro
10 Apr 2026
Research

Iteratively Learning Muscle Memory for Legged Robots to Master Adaptive and High Precision Locomotion

DGX agent

arXiv:2507.13662v2 Announce Type: replace Abstract: This paper presents a scalable and adaptive control framework for legged robots that integrates Iterative Learning Control (ILC) with a biologically

researcharxiv-cs-ro
10 Apr 2026
Research

LongSpec: Long-Context Lossless Speculative Decoding with Efficient Drafting and Verification

DGX agent

arXiv:2502.17421v4 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) can now process extremely long contexts, efficient inference over these extended inputs has become increasingl

researcharxiv-cs-ai
10 Apr 2026
Model Releases

LongWriter-Zero: Mastering Ultra-Long Text Generation via Reinforcement Learning

DGX agent

arXiv:2506.18841v3 Announce Type: replace-cross Abstract: Ultra-long generation by large language models (LLMs) is a widely demanded scenario, yet it remains a significant challenge due to their maxim

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

LPM 1.0: Video-based Character Performance Model

DGX agent

arXiv:2604.07823v1 Announce Type: new Abstract: Performance, the externalization of intent, emotion, and personality through visual, vocal, and temporal behavior, is what makes a character alive. Lear

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

MARCH: Evaluating the Intersection of Ambiguity Interpretation and Multi-hop Inference

DGX agent

arXiv:2509.22750v3 Announce Type: replace Abstract: Real-world multi-hop QA is naturally linked with ambiguity, where a single query can trigger multiple reasoning paths that require independent resol

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Mitigating Distribution Sharpening in Math RLVR via Distribution-Aligned Hint Synthesis and Backward Hint Annealing

DGX agent

arXiv:2604.07747v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) can improve low-k reasoning accuracy while narrowing solution coverage on challenging math que

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

MorphDistill: Distilling Unified Morphological Knowledge from Pathology Foundation Models for Colorectal Cancer Survival Prediction

DGX agent

arXiv:2604.06390v1 Announce Type: cross Abstract: Background: Colorectal cancer (CRC) remains a leading cause of cancer-related mortality worldwide. Accurate survival prediction is essential for treat

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Orion-Lite: Distilling LLM Reasoning into Efficient Vision-Only Driving Models

DGX agent

arXiv:2604.08266v1 Announce Type: new Abstract: Leveraging the general world knowledge of Large Language Models (LLMs) holds significant promise for improving the ability of autonomous driving systems

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

PokeGym: A Visually-Driven Long-Horizon Benchmark for Vision-Language Models

DGX agent

arXiv:2604.08340v1 Announce Type: new Abstract: While Vision-Language Models (VLMs) have achieved remarkable progress in static visual understanding, their deployment in complex 3D embodied environmen

model-releasesarxiv-cs-cv
10 Apr 2026
Local Ai

Q-Zoom: Query-Aware Adaptive Perception for Efficient Multimodal Large Language Models

DGX agent

arXiv:2604.06912v1 Announce Type: cross Abstract: MLLMs require high-resolution visual inputs for fine-grained tasks like document understanding and dense scene perception. However, current global res

local-aiarxiv-cs-ai
10 Apr 2026
Tutorials

Reading Recognition in the Wild

DGX agent

arXiv:2505.24848v4 Announce Type: replace Abstract: To enable egocentric contextual AI in always-on smart glasses, it is crucial to be able to keep a record of the user's interactions with the world,

tutorialsarxiv-cs-cv
10 Apr 2026
Safety

Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning

DGX agent

arXiv:2505.24499v2 Announce Type: replace Abstract: Generating high-quality Scalable Vector Graphics (SVGs) is challenging for Large Language Models (LLMs), as it requires advanced reasoning for struc

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

Riemann-Bench: A Benchmark for Moonshot Mathematics

DGX agent

arXiv:2604.06802v1 Announce Type: new Abstract: Recent AI systems have achieved gold-medal-level performance on the International Mathematical Olympiad, demonstrating remarkable proficiency at competi

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

SciFigDetect: A Benchmark for AI-Generated Scientific Figure Detection

DGX agent

arXiv:2604.08211v1 Announce Type: new Abstract: Modern multimodal generators can now produce scientific figures at near-publishable quality, creating a new challenge for visual forensics and research

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

sciwrite-lint: Verification Infrastructure for the Age of Science Vibe-Writing

DGX agent

arXiv:2604.08501v1 Announce Type: cross Abstract: Science currently offers two options for quality assurance, both inadequate. Journal gatekeeping claims to verify both integrity and contribution, but

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SealQA: Raising the Bar for Reasoning in Search-Augmented Language Models

DGX agent

arXiv:2506.01062v4 Announce Type: replace Abstract: We introduce SealQA, a new challenge benchmark for evaluating SEarch-Augmented Language models on fact-seeking questions where web search yields con

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Self-Distilled RLVR

DGX agent

arXiv:2604.03128v2 Announce Type: replace Abstract: On-policy distillation (OPD) has become a popular training paradigm in the LLM community. This paradigm selects a larger model as the teacher to pro

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

Sell More, Play Less: Benchmarking LLM Realistic Selling Skill

DGX agent

arXiv:2604.07054v2 Announce Type: replace Abstract: Sales dialogues require multi-turn, goal-directed persuasion under asymmetric incentives, which makes them a challenging setting for large language

model-releasesarxiv-cs-cl
10 Apr 2026
Local Ai

SepSeq: A Training-Free Framework for Long Numerical Sequence Processing in LLMs

DGX agent

arXiv:2604.07737v1 Announce Type: new Abstract: While transformer-based Large Language Models (LLMs) theoretically support massive context windows, they suffer from severe performance degradation when

local-aiarxiv-cs-cl
10 Apr 2026
← Previous
1…233234235236
Next →