AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Model Releases

Language Models Don't Know What You Want: Evaluating Personalization in Deep Research Needs Real Users

DGX agent

arXiv:2603.16120v2 Announce Type: replace Abstract: Deep Research (DR) systems help researchers cope with ballooning publishing counts. Such tools synthesize scientific papers to answer research queri

model-releasesarxiv-cs-cl
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Learning-Based Sparsification of Dynamic Graphs in Robotic Exploration Algorithms

DGX agent

arXiv:2604.16509v1 Announce Type: cross Abstract: Many robotic exploration algorithms rely on graph structures for frontier-based exploration and dynamic path planning. However, these graphs grow rapi

safetyarxiv-cs-lg
21 Apr 2026
Safety

MHSafeEval: Role-Aware Interaction-Level Evaluation of Mental Health Safety in Large Language Models

DGX agent

arXiv:2604.17730v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly explored as scalable tools for mental health counseling, yet evaluating their safety remains challenging d

safetyarxiv-cs-cl
21 Apr 2026
Safety

Modeling User Exploration Saturation: When Recommender Systems Should Stop Pushing Novelty

DGX agent

arXiv:2604.16419v1 Announce Type: cross Abstract: Fairness-aware recommender systems often mitigate bias by increasing exposure to under-represented or long-tail content, commonly through mechanisms t

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Neuro-Symbolic Resolution of Recommendation Conflicts in Multimorbidity Clinical Guidelines

DGX agent

arXiv:2604.17340v1 Announce Type: new Abstract: Clinical guidelines, typically developed by independent specialty societies, inherently exhibit substantial fragmentation, redundancy, and logical contr

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

NL2SQLBench: A Modular Benchmarking Framework for LLM-Enabled NL2SQL Solutions

DGX agent

arXiv:2604.16493v1 Announce Type: cross Abstract: Natural Language to SQL (NL2SQL) technology empowers non-expert users to query relational databases without requiring SQL expertise. While large langu

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

OPeRA: A Dataset of Observation, Persona, Rationale, and Action for Evaluating LLMs on Human Online Shopping Behavior Simulation

DGX agent

arXiv:2506.05606v5 Announce Type: replace Abstract: Can large language models (LLMs) accurately simulate the next web action of a specific user? While LLMs have shown promising capabilities in generat

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Plasticity Loss in Deep Reinforcement Learning: A Survey

DGX agent

arXiv:2411.04832v3 Announce Type: replace-cross Abstract: Plasticity refers to a network's ability to adapt to changing data distributions, which is crucial for the successful training of deep reinfor

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Precise Debugging Benchmark: Is Your Model Debugging or Regenerating?

DGX agent

arXiv:2604.17338v1 Announce Type: cross Abstract: Unlike code completion, debugging requires localizing faults and applying targeted edits. We observe that frontier LLMs often regenerate correct but o

model-releasesarxiv-cs-cl
21 Apr 2026
Research

PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues

DGX agent

arXiv:2604.18354v1 Announce Type: new Abstract: Emotion plays a pivotal role in shaping negotiation outcomes, influencing trust, cooperation, and long-term relationships. Developing negotiation dialog

researcharxiv-cs-cl
21 Apr 2026
Safety

R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation

DGX agent

arXiv:2506.07826v2 Announce Type: replace Abstract: Validating autonomous driving (AD) systems requires diverse and safety-critical testing, making photorealistic virtual environments essential. Tradi

safetyarxiv-cs-cv
21 Apr 2026
Research

REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations

DGX agent

arXiv:2604.17289v1 Announce Type: new Abstract: Supervised fine-tuning of large language models relies on human-annotated data, yet annotation pipelines routinely involve multiple crowdworkers of hete

researcharxiv-cs-lg
21 Apr 2026
Model Releases

Scaling Human-AI Coding Collaboration Requires a Governable Consensus Layer

DGX agent

arXiv:2604.17883v1 Announce Type: cross Abstract: Vibe coding produces correct, executable code at speed, but leaves no record of the structural commitments, dependencies, or evidence behind it. Revie

model-releasesarxiv-cs-lg
21 Apr 2026
Research

ScenarioControl: Vision-Language Controllable Vectorized Latent Scenario Generation

DGX agent

arXiv:2604.17147v1 Announce Type: new Abstract: We introduce ScenarioControl, the first vision-language control mechanism for learned driving scenario generation. Given a text prompt or an input image

researcharxiv-cs-cv
21 Apr 2026
Research

SentiAvatar: Towards Expressive and Interactive Digital Humans

DGX agent

arXiv:2604.02908v2 Announce Type: replace Abstract: We present SentiAvatar, a framework for building expressive interactive 3D digital humans, and use it to create SuSu, a virtual character that speak

researcharxiv-cs-cv
21 Apr 2026
Local Ai

ST-pi: Structured SpatioTemporal VLA for Robotic Manipulation

DGX agent

arXiv:2604.17880v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have achieved great success on general robotic tasks, but still face challenges in fine-grained spatiotemporal man

local-aiarxiv-cs-cv
21 Apr 2026
Research

Stable Language Guidance for Vision-Language-Action Models

DGX agent

arXiv:2601.04052v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models have demonstrated impressive capabilities in generalized robotic control; however, they remain notoriously

researcharxiv-cs-cl
21 Apr 2026
Research

SVL: Goal-Conditioned Reinforcement Learning as Survival Learning

DGX agent

arXiv:2604.17551v1 Announce Type: new Abstract: Standard approaches to goal-conditioned reinforcement learning (GCRL) that rely on temporal-difference learning can be unstable and sample-inefficient d

researcharxiv-cs-lg
21 Apr 2026
Safety

Synthia: Scalable Grounded Persona Generation from Social Media Data

DGX agent

arXiv:2507.14922v2 Announce Type: replace Abstract: Persona-driven simulations are increasingly used in computational social science, yet their validity critically depends on the fidelity of the under

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

TeleEmbedBench: A Multi-Corpus Embedding Benchmark for RAG in Telecommunications

DGX agent

arXiv:2604.17778v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in the telecommunications domain for critical tasks, relying heavily on Retrieval-Augmented Gener

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models

DGX agent

arXiv:2604.18107v1 Announce Type: new Abstract: Vision-Language-Action models (VLAs) achieve remarkable performance in sequential decision-making but remain fragile to subtle environmental shifts, suc

researcharxiv-cs-cv
21 Apr 2026
Research

Towards Disentangled Preference Optimization Dynamics Beyond Likelihood Displacement

DGX agent

arXiv:2604.18239v1 Announce Type: new Abstract: Preference optimization is widely used to align large language models (LLMs) with human preferences. However, many margin-based objectives suppress the

researcharxiv-cs-lg
21 Apr 2026
Model Releases

TSVer: A Benchmark for Fact Verification Against Time-Series Evidence

DGX agent

arXiv:2511.01101v2 Announce Type: replace Abstract: Reasoning over temporal and numerical data, such as time series, is a crucial aspect of fact-checking. While many systems have recently been develop

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Untrained CNNs Match Backpropagation at V1: A Systematic RSA Comparison of Four Learning Rules Against Human fMRI

DGX agent

arXiv:2604.16875v1 Announce Type: new Abstract: A central question in computational neuroscience is whether the learning rule used to train a neural network determines how well its internal representa

safetyarxiv-cs-lg
21 Apr 2026
Research

Using Perspectival Words Is Harder Than Vocabulary Words for Humans and Even More So for Multimodal Language Models

DGX agent

arXiv:2506.00065v2 Announce Type: replace Abstract: Multimodal language models (MLMs) increasingly demonstrate human-like communication, yet their use of everyday perspectival words remains poorly und

researcharxiv-cs-cl
21 Apr 2026
Safety

Where Do Self-Supervised Speech Models Become Unfair?

DGX agent

arXiv:2604.18249v1 Announce Type: new Abstract: Speech encoder models are known to model members of some speaker groups (SGs) better than others. However, there has been little work in establishing wh

safetyarxiv-cs-cl
21 Apr 2026
Local Ai

AdaVFM: Adaptive Vision Foundation Models for Edge Intelligence via LLM-Guided Execution

DGX agent

arXiv:2604.15622v1 Announce Type: new Abstract: Language-aligned vision foundation models (VFMs) enable versatile visual understanding for always-on contextual AI, but their deployment on edge devices

local-aiarxiv-cs-cv
20 Apr 2026
Research

Anthropomorphism and Trust in Human-Large Language Model interactions

DGX agent

arXiv:2604.15316v1 Announce Type: cross Abstract: With large language models (LLMs) becoming increasingly prevalent in daily life, so too has the tendency to attribute to them human-like minds and emo

researcharxiv-cs-ai
20 Apr 2026
Model Releases

AutoFed: Personalized Federated Traffic Prediction via Adaptive Prompt

DGX agent

arXiv:2512.24625v2 Announce Type: replace-cross Abstract: Accurate traffic prediction is essential for Intelligent Transportation Systems, including ride-hailing, urban road planning, and vehicle flee

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Beyond Distribution Sharpening: The Importance of Task Rewards

DGX agent

arXiv:2604.16259v1 Announce Type: cross Abstract: Frontier models have demonstrated exceptional capabilities following the integration of task-reward-based reinforcement learning (RL) into their train

model-releasesarxiv-cs-ai
20 Apr 2026
Local Ai

Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning

DGX agent

arXiv:2604.15414v1 Announce Type: cross Abstract: Continual reinforcement learning must balance retention with adaptation, yet many methods still rely on single-model preservation, committing to one e

local-aiarxiv-cs-ai
20 Apr 2026
Tutorials

Bridging the phenotype-target gap for molecular generation via multi-objective reinforcement learning

DGX agent

arXiv:2509.21010v2 Announce Type: replace-cross Abstract: The de novo generation of drug-like molecules capable of inducing desirable phenotypic changes is receiving increasing attention. However, pre

tutorialsarxiv-cs-ai
20 Apr 2026
Model Releases

CoMeT: Collaborative Memory Transformer for Efficient Long Context Modeling

DGX agent

arXiv:2602.01766v2 Announce Type: replace-cross Abstract: The quadratic complexity and indefinitely growing key-value (KV) cache of standard Transformers pose a major barrier to long-context processin

model-releasesarxiv-cs-ai
20 Apr 2026
Local Ai

Continual Hand-Eye Calibration for Open-world Robotic Manipulation

DGX agent

arXiv:2604.15814v1 Announce Type: new Abstract: Hand-eye calibration through visual localization is a critical capability for robotic manipulation in open-world environments. However, most deep learni

local-aiarxiv-cs-cv
20 Apr 2026
Safety

Deliberative Searcher: Improving LLM Reliability via Reinforcement Learning with constraints

DGX agent

arXiv:2507.16727v3 Announce Type: replace Abstract: Improving the reliability of large language models (LLMs) is critical for deploying them in real-world scenarios. In this paper, we propose extbf{De

safetyarxiv-cs-ai
20 Apr 2026
Safety

Exploitation Over Exploration: Unmasking the Bias in Linear Bandit Recommender Offline Evaluation

DGX agent

arXiv:2507.18756v2 Announce Type: replace Abstract: Multi-Armed Bandit (MAB) algorithms are widely used in recommender systems that require continuous, incremental learning. A core aspect of MABs is t

safetyarxiv-cs-lg
20 Apr 2026
Model Releases

Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation

DGX agent

arXiv:2604.15388v1 Announce Type: cross Abstract: Recent advances in large language models have improved code generation, but their use in hardware description languages is still limited. Moreover, tr

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Foundation Models in Robotics: A Comprehensive Review of Methods, Models, Datasets, Challenges and Future Research Directions

DGX agent

arXiv:2604.15395v1 Announce Type: new Abstract: Over the recent years, the field of robotics has been undergoing a transformative paradigm shift from fixed, single-task, domain-specific solutions towa

applicationsarxiv-cs-ro
20 Apr 2026
Safety

From Intention to Text: AI-Supported Goal Setting in Academic Writing

DGX agent

arXiv:2604.15800v1 Announce Type: cross Abstract: This study presents WriteFlow, an AI voice-based writing assistant designed to support reflective academic writing through goal-oriented interaction.

safetyarxiv-cs-ai
20 Apr 2026
Safety

How people use Copilot for Health

DGX agent

arXiv:2604.15331v1 Announce Type: cross Abstract: We analyze over 500,000 de-identified health-related conversations with Microsoft Copilot from January 2026 to characterize what people ask conversati

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

KWBench: Measuring Unprompted Problem Recognition in Knowledge Work

DGX agent

arXiv:2604.15760v1 Announce Type: new Abstract: We introduce the first version of KWBench (Knowledge Work Bench), a benchmark for unprompted problem recognition in large language models: can an LLM id

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

LLMs Corrupt Your Documents When You Delegate

DGX agent

arXiv:2604.15597v1 Announce Type: new Abstract: Large Language Models (LLMs) are poised to disrupt knowledge work, with the emergence of delegated work as a new interaction paradigm (e.g., vibe coding

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Neurosymbolic Repo-level Code Localization

DGX agent

arXiv:2604.16021v1 Announce Type: cross Abstract: Code localization is a cornerstone of autonomous software engineering. Recent advancements have achieved impressive performance on real-world issue be

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

PAWN: Piece Value Analysis with Neural Networks

DGX agent

arXiv:2604.15585v1 Announce Type: cross Abstract: Predicting the relative value of any given chess piece in a position remains an open challenge, as a piece's contribution depends on its spatial relat

safetyarxiv-cs-ai
20 Apr 2026
Research

Placing Puzzle Pieces Where They Matter: A Question Augmentation Framework for Reinforcement Learning

DGX agent

arXiv:2604.15830v1 Announce Type: new Abstract: Reinforcement learning has become a powerful approach for enhancing large language model reasoning, but faces a fundamental dilemma: training on easy pr

researcharxiv-cs-lg
20 Apr 2026
Safety

Puppets or partners? Governing cyborg propaganda in the digital public square

DGX agent

arXiv:2602.13088v2 Announce Type: replace-cross Abstract: The distinction between genuine grassroots activism and automated influence operations is collapsing. While contemporary policy debates priori

safetyarxiv-cs-ai
20 Apr 2026
Safety

Safe and Energy-Aware Multi-Robot Density Control via PDE-Constrained Optimization for Long-Duration Autonomy

DGX agent

arXiv:2604.15524v1 Announce Type: cross Abstract: This paper presents a novel density control framework for multi-robot systems with spatial safety and energy sustainability guarantees. Stochastic rob

safetyarxiv-cs-ro
20 Apr 2026
Safety

Safe Deep Reinforcement Learning for Building Heating Control and Demand-side Flexibility

DGX agent

arXiv:2604.16033v1 Announce Type: cross Abstract: Buildings account for approximately 40% of global energy consumption, and with the growing share of intermittent renewable energy sources, enabling de

safetyarxiv-cs-ai
20 Apr 2026
← Previous
1…224225226227228…230
Next →