AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Local Ai

Learning Reasoning Reward Models from Expert Demonstration via Inverse Reinforcement Learning

DGX agent

arXiv:2510.01857v3 Announce Type: replace Abstract: Current approaches to improving reasoning in large language models (LLMs) primarily rely on either supervised fine-tuning (SFT) over expert traces o

local-aiarxiv-cs-ai
24 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Task-specific Subnetwork Discovery in Reinforcement Learning for Autonomous Underwater Navigation

DGX agent

arXiv:2604.21640v1 Announce Type: cross Abstract: Autonomous underwater vehicles are required to perform multiple tasks adaptively and in an explainable manner under dynamic, uncertain conditions and

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

SweRank: Software Issue Localization with Code Ranking

DGX agent

arXiv:2505.07849v2 Announce Type: replace-cross Abstract: Software issue localization, the task of identifying the precise code locations (files, classes, or functions) relevant to a natural language

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Mitigating Judgment Preference Bias in Large Language Models through Group-Based Polling

DGX agent

arXiv:2510.08145v2 Announce Type: replace Abstract: Large Language Models (LLMs) as automatic evaluators, commonly referred to as LLM-as-a-Judge, have also attracted growing attention. This approach p

safetyarxiv-cs-cl
22 Apr 2026
Safety

Decoding AI Tutor Effects for Educational Measurement: Temporal, Multi-Outcome, and Behavior-Cognitive Analysis

DGX agent

arXiv:2604.16366v1 Announce Type: cross Abstract: Artificial intelligence (AI) tutors have become increasingly popular in learning environments. In this study, we propose an AI agent prototype framewo

safetyarxiv-cs-lg
21 Apr 2026
Safety

Heterogeneous Self-Play for Realistic Highway Traffic Simulation

DGX agent

arXiv:2604.16406v1 Announce Type: cross Abstract: Realistic highway simulation is critical for scalable safety evaluation of autonomous vehicles, particularly for interactions that are too rare to stu

safetyarxiv-cs-lg
21 Apr 2026
Agents

LookasideVLN: Direction-Aware Aerial Vision-and-Language Navigation

DGX agent

arXiv:2604.17190v1 Announce Type: new Abstract: Aerial Vision-and-Language Navigation (Aerial VLN) enables unmanned aerial vehicles (UAVs) to follow natural language instructions and navigate complex

agentsarxiv-cs-cv
21 Apr 2026
Safety

MoCo: A One-Stop Shop for Model Collaboration Research

DGX agent

arXiv:2601.21257v2 Announce Type: replace Abstract: Advancing beyond single monolithic language models (LMs), recent research increasingly recognizes the importance of model collaboration, where multi

safetyarxiv-cs-cl
21 Apr 2026
Agents

OneVL: One-Step Latent Reasoning and Planning with Vision-Language Explanation

DGX agent

arXiv:2604.18486v1 Announce Type: cross Abstract: Chain-of-Thought (CoT) reasoning has become a powerful driver of trajectory prediction in VLA-based autonomous driving, yet its autoregressive nature

agentsarxiv-cs-cl
21 Apr 2026
Research

Tool Learning Needs Nothing More Than a Free 8B Language Model

DGX agent

arXiv:2604.17739v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a prevalent paradigm for training tool calling agents, which typically requires online interactive environments

researcharxiv-cs-cl
21 Apr 2026
Safety

Aerial Multi-Functional RIS in Fluid Antennas-Aided Full-Duplex Networks: A Self-Optimized Hybrid Deep Reinforcement Learning Approach

DGX agent

arXiv:2604.14309v2 Announce Type: replace-cross Abstract: To address high data traffic demands of sixth-generation (6G) networks, this paper proposes a novel architecture that integrates autonomous ae

safetyarxiv-cs-ai
20 Apr 2026
Safety

NLP needs Diversity outside of 'Diversity'

DGX agent

arXiv:2604.14595v1 Announce Type: new Abstract: This position paper argues that recent progress with diversity in NLP is disproportionately concentrated on a small number of areas surrounding fairness

safetyarxiv-cs-cl
17 Apr 2026
Agents

TableNet A Large-Scale Table Dataset with LLM-Powered Autonomous

DGX agent

arXiv:2604.13041v1 Announce Type: cross Abstract: Table Structure Recognition (TSR) requires the logical reasoning ability of large language models (LLMs) to handle complex table layouts, but current

agentsarxiv-cs-ai
17 Apr 2026
Tutorials

World-Value-Action Model: Implicit Planning for Vision-Language-Action Systems

DGX agent

arXiv:2604.14732v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a promising paradigm for building embodied agents that ground perception and language into action.

tutorialsarxiv-cs-lg
17 Apr 2026
Safety

Beyond State Consistency: Behavior Consistency in Text-Based World Models

DGX agent

arXiv:2604.13824v1 Announce Type: new Abstract: World models have been emerging as critical components for assessing the consequences of actions generated by interactive agents in online planning and

safetyarxiv-cs-lg
16 Apr 2026
Safety

C^2T: Captioning-Structure and LLM-Aligned Common-Sense Reward Learning for Traffic--Vehicle Coordination

DGX agent

arXiv:2604.13098v1 Announce Type: cross Abstract: State-of-the-art (SOTA) urban traffic control increasingly employs Multi-Agent Reinforcement Learning (MARL) to coordinate Traffic Light Controllers (

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

Red Skills or Blue Skills? A Dive Into Skills Published on ClawHub

DGX agent

arXiv:2604.13064v1 Announce Type: new Abstract: Skill ecosystems have emerged as an increasingly important layer in Large Language Model (LLM) agent systems, enabling reusable task packaging, public d

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

Training-Free Test-Time Contrastive Learning for Large Language Models

DGX agent

arXiv:2604.13552v1 Announce Type: new Abstract: Large language models (LLMs) demonstrate strong reasoning capabilities, but their performance often degrades under distribution shift. Existing test-tim

agentsarxiv-cs-cl
16 Apr 2026
Safety

A Comparison of Reinforcement Learning and Optimal Control Methods for Path Planning

DGX agent

arXiv:2604.12628v1 Announce Type: cross Abstract: Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge. While traditional optimal control methods can find idea

safetyarxiv-cs-ro
15 Apr 2026
Agents

BLOSSOM: Block-wise Federated Learning Over Shared and Sparse Observed Modalities

DGX agent

arXiv:2603.27552v2 Announce Type: replace Abstract: Multimodal federated learning (FL) is essential for real-world applications such as autonomous systems and healthcare, where data is distributed acr

agentsarxiv-cs-lg
15 Apr 2026
Model Releases

CompliBench: Benchmarking LLM Judges for Compliance Violation Detection in Dialogue Systems

DGX agent

arXiv:2604.12312v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as task-oriented agents in enterprise environments, ensuring their strict adherence to complex

model-releasesarxiv-cs-cl
15 Apr 2026
Safety

Incentivizing High-Quality Human Annotations with Golden Questions

DGX agent

arXiv:2505.19134v2 Announce Type: replace-cross Abstract: Human-annotated data plays a vital role in training large language models (LLMs), such as supervised fine-tuning and human preference alignmen

safetyarxiv-cs-lg
15 Apr 2026
Agents

Thermodynamic Liquid Manifold Networks: Physics-Bounded Deep Learning for Solar Forecasting in Autonomous Off-Grid Microgrids

DGX agent

arXiv:2604.11909v1 Announce Type: cross Abstract: The stable operation of autonomous off-grid photovoltaic systems requires solar forecasting algorithms that respect atmospheric thermodynamics. Contem

agentsarxiv-cs-ai
15 Apr 2026
Agents

A Survey on Deep Learning Techniques for Action Anticipation

DGX agent

arXiv:2309.17257v2 Announce Type: replace Abstract: The ability to anticipate possible future human actions is essential for a wide range of applications, including autonomous driving and human-robot

agentsarxiv-cs-cv
14 Apr 2026
Model Releases

ActorMind: Emulating Human Actor Reasoning for Speech Role-Playing

DGX agent

arXiv:2604.11103v1 Announce Type: cross Abstract: Role-playing has garnered rising attention as it provides a strong foundation for human-machine interaction and facilitates sociological research. How

model-releasesarxiv-cs-ai
14 Apr 2026
Hardware

Characterizing Performance-Energy Trade-offs of Large Language Models in Multi-Request Workflows

DGX agent

arXiv:2604.09611v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in applications forming multi-request workflows like document summarization, search-based copilots,

hardwarearxiv-cs-ai
14 Apr 2026
Model Releases

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

DGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

Energy-Efficient Federated Edge Learning For Small-Scale Datasets in Large IoT Networks

DGX agent

arXiv:2604.10662v1 Announce Type: new Abstract: Large-scale Internet of Things (IoT) networks enable intelligent services such as smart cities and autonomous driving, but often face resource constrain

agentsarxiv-cs-lg
14 Apr 2026
Model Releases

GroupRank: A Groupwise Paradigm for Effective and Efficient Passage Reranking with LLMs

DGX agent

arXiv:2511.11653v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have emerged as powerful tools for passage reranking in information retrieval, leveraging their superior reasonin

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

If an LLM Were a Character, Would It Know Its Own Story? Evaluating Lifelong Learning in LLMs

DGX agent

arXiv:2503.23514v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can carry out human-like dialogue, but unlike humans, they are stateless due to the superposition property. Howev

model-releasesarxiv-cs-ai
14 Apr 2026
Local Ai

MeloTune: On-Device Arousal Learning and Peer-to-Peer Mood Coupling for Proactive Music Curation

DGX agent

arXiv:2604.10815v1 Announce Type: cross Abstract: MeloTune is an iPhone-deployed music agent that instantiates the Mesh Memory Protocol (MMP) and Symbolic-Vector Attention Fusion (SVAF) as a productio

local-aiarxiv-cs-ai
14 Apr 2026
Safety

Rebooting Microreboot: Architectural Support for Safe, Parallel Recovery in Microservice Systems

DGX agent

arXiv:2604.09963v1 Announce Type: cross Abstract: Microreboot enables fast recovery by restarting only the failing component, but in modern microservices naive restarts are unsafe: dense dependencies

safetyarxiv-cs-ai
14 Apr 2026
Safety

Shuffling the Data, Stretching the Step-size: Sharper Bias in constant step-size SGD

DGX agent

arXiv:2604.10373v1 Announce Type: cross Abstract: From adversarial robustness to multi-agent learning, many machine learning tasks can be cast as finite-sum min-max optimization or, more generally, as

safetyarxiv-cs-lg
14 Apr 2026
Agents

Unified Unsupervised and Sparsely-Supervised 3D Object Detection by Semantic Pseudo-Labeling and Prototype Learning

DGX agent

arXiv:2602.21484v2 Announce Type: replace Abstract: 3D object detection is essential for autonomous driving and robotic perception, yet its reliance on large-scale manually annotated data limits scala

agentsarxiv-cs-cv
14 Apr 2026
Safety

Utilizing and Calibrating Hindsight Process Rewards via Reinforcement with Mutual Information Self-Evaluation

DGX agent

arXiv:2604.11611v1 Announce Type: new Abstract: To overcome the sparse reward challenge in reinforcement learning (RL) for agents based on large language models (LLMs), we propose Mutual Information S

safetyarxiv-cs-cl
14 Apr 2026
Hardware

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference

DGX agent

arXiv:2604.08584v1 Announce Type: cross Abstract: Long-context LLMs increasingly rely on extended, reusable prefill prompts for agents and domain Q&A, pushing attention and KV-cache to become the domi

hardwarearxiv-cs-ai
13 Apr 2026
Agents

Decentralized Opinion-Integrated Decision making at Unsignalized Intersections via Signed Networks

DGX agent

arXiv:2604.09351v1 Announce Type: cross Abstract: In this letter, we consider the problem of decentralized decision making among connected autonomous vehicles at unsignalized intersections, where exis

agentsarxiv-cs-ro
13 Apr 2026
Agents

Koopman Operator Framework for Modeling and Control of Off-Road Vehicle on Deformable Terrain

DGX agent

arXiv:2603.28965v2 Announce Type: replace-cross Abstract: This work presents a hybrid physics-informed and data-driven modeling framework for predictive control of autonomous off-road vehicles operati

agentsarxiv-cs-ro
13 Apr 2026
Agents

Yes, But Not Always. Generative AI Needs Nuanced Opt-in

DGX agent

arXiv:2604.09413v1 Announce Type: cross Abstract: This paper argues that a one-size-fits-all approach to specifying consent for the use of creative works in generative AI is insufficient. Real-world o

agentsarxiv-cs-ai
13 Apr 2026
Agents

Assessing the Added Value of Onboard Earth Observation Processing with the IRIDE HEO Service Segment

DGX agent

arXiv:2604.07120v1 Announce Type: cross Abstract: Current operational Earth Observation (EO) services, including the Copernicus Emergency Management Service (CEMS), the European Forest Fire Informatio

agentsarxiv-cs-ai
10 Apr 2026
Agents

Formally Guaranteed Control Adaptation for ODD-Resilient Autonomous Systems

DGX agent

arXiv:2604.07414v1 Announce Type: cross Abstract: Ensuring reliable performance in situations outside the Operational Design Domain (ODD) remains a primary challenge in devising resilient autonomous s

agentsarxiv-cs-ro
10 Apr 2026
Model Releases

HyperMem: Hypergraph Memory for Long-Term Conversations

DGX agent

arXiv:2604.08256v1 Announce Type: new Abstract: Long-term memory is essential for conversational agents to maintain coherence, track persistent tasks, and provide personalized interactions across exte

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

Semantic-Aware UAV Command and Control for Efficient IoT Data Collection

DGX agent

arXiv:2604.08153v1 Announce Type: new Abstract: Unmanned Aerial Vehicles (UAVs) have emerged as a key enabler technology for data collection from Internet of Things (IoT) devices. However, effective d

safetyarxiv-cs-ro
10 Apr 2026
Safety

The End of the Foundation Model Era: Open-Weight Models, Sovereign AI, and Inference as Infrastructure

DGX agent

arXiv:2604.06217v1 Announce Type: cross Abstract: The foundation model era -- roughly 2020 to 2025 -- is over. The forces that defined it have inverted. Open source models have reached frontier perfor

safetyarxiv-cs-ai
10 Apr 2026
Agents

AMR-Pose: An Active LED Marker-Based Relative Pose Estimation Framework With Probabilistic Switching PnP for Cooperative AUVs

DGX agent

arXiv:2608.12866v1 Announce Type: new Abstract: Reliable relative pose estimation between autonomous underwater vehicles (AUVs) is critical for cooperative ocean exploration, sampling, and multi-robot

agentsarxiv-cs-ro
14 Aug 2026
Agents

BrainWAM: Action-Space Coordination of Semantic Priors and Predictive Dynamics for Autonomous Driving

DGX agent

arXiv:2608.12854v1 Announce Type: cross Abstract: Autonomous driving requires planning under both semantic constraints and predictive dynamics. Existing end-to-end driving approaches, however, typical

agentsarxiv-cs-ai
14 Aug 2026
Safety

Entropy-Augmented Multi-Objective Policy Optimization in Multiagent Systems

DGX agent

arXiv:2608.12534v1 Announce Type: cross Abstract: Autonomous agent teams deployed in settings such as marine and extraterrestrial outposts must coordinate actions to achieve optimal outcomes across mu

safetyarxiv-cs-ro
14 Aug 2026
Model Releases

FUSE: Active Functional Affordance Grounding through Adaptive Semantic-Geometric Evidence Acquisition

DGX agent

arXiv:2608.12683v1 Announce Type: cross Abstract: Embodied agents must often identify and interact with objects based on their function rather than their identity, requiring them to actively acquire o

model-releasesarxiv-cs-cv
14 Aug 2026
← Previous
1…150151152153154…236
Next →