AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

Beyond Pessimism: Offline Learning in KL-regularized Games

DGX agent

arXiv:2604.06738v1 Announce Type: cross Abstract: We study offline learning in KL-regularized two-player zero-sum games, where policies are optimized under a KL constraint to a fixed reference policy.

safetyarxiv-cs-lg
10 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Brain3D: EEG-to-3D Decoding of Visual Representations via Multimodal Reasoning

DGX agent

arXiv:2604.08068v1 Announce Type: new Abstract: Decoding visual information from electroencephalography (EEG) has recently achieved promising results, primarily focusing on reconstructing two-dimensio

safetyarxiv-cs-cv
10 Apr 2026
Local Ai

Chatbot-Based Assessment of Code Understanding in Automated Programming Assessment Systems

DGX agent

arXiv:2604.07304v1 Announce Type: cross Abstract: Large Language Models (LLMs) challenge conventional automated programming assessment because students can now produce functionally correct code withou

local-aiarxiv-cs-ai
10 Apr 2026
Hardware

CodecSight: Leveraging Video Codec Signals for Efficient Streaming VLM Inference

DGX agent

arXiv:2604.06036v3 Announce Type: replace-cross Abstract: Video streaming analytics is a crucial workload for vision-language model serving, but the high cost of multimodal inference limits scalabilit

hardwarearxiv-cs-cv
10 Apr 2026
Model Releases

ConvoLearn: A Dataset for Fine-Tuning Dialogic AI Tutors

DGX agent

arXiv:2601.08950v3 Announce Type: replace Abstract: Despite their growing adoption in education, LLMs remain misaligned with the core principle of effective tutoring: the dialogic construction of know

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Data Leakage in Automotive Perception: Practitioners' Insights

DGX agent

arXiv:2604.06899v1 Announce Type: cross Abstract: Data leakage is the inadvertent transfer of information between training and evaluation datasets that poses a subtle, yet critical, risk to the reliab

safetyarxiv-cs-lg
10 Apr 2026
Safety

Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs

DGX agent

arXiv:2604.07518v1 Announce Type: new Abstract: Vision-Language Models often struggle with complex visual reasoning due to the visual information loss in textual CoT. Existing methods either add the c

safetyarxiv-cs-cl
10 Apr 2026
Safety

Demystifying OPD: Length Inflation and Stabilization Strategies for Large Language Models

DGX agent

arXiv:2604.08527v1 Announce Type: new Abstract: On-policy distillation (OPD) trains student models under their own induced distribution while leveraging supervision from stronger teachers. We identify

safetyarxiv-cs-cl
10 Apr 2026
Safety

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

DGX agent

arXiv:2604.07607v1 Announce Type: cross Abstract: Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human da

safetyarxiv-cs-cv
10 Apr 2026
Hardware

Foundry: Template-Based CUDA Graph Context Materialization for Fast LLM Serving Cold Start

DGX agent

arXiv:2604.06664v1 Announce Type: cross Abstract: Modern LLM service providers increasingly rely on autoscaling and parallelism reconfiguration to respond to rapidly changing workloads, but cold-start

hardwarearxiv-cs-lg
10 Apr 2026
Safety

FP4 Explore, BF16 Train: Diffusion Reinforcement Learning via Efficient Rollout Scaling

DGX agent

arXiv:2604.06916v1 Announce Type: cross Abstract: Reinforcement-Learning-based post-training has recently emerged as a promising paradigm for aligning text-to-image diffusion models with human prefere

safetyarxiv-cs-ai
10 Apr 2026
Safety

How Psychological Learning Paradigms Shaped and Constrained Artificial Intelligence

DGX agent

arXiv:2603.18203v3 Announce Type: replace Abstract: Current artificial intelligence systems struggle with systematic compositional reasoning: the capacity to recombine known components in novel config

safetyarxiv-cs-cl
10 Apr 2026
Safety

HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns

DGX agent

arXiv:2601.10198v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and generation, serving as the foundation for advanced persona s

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

k-server-bench: Automating Potential Discovery for the k-Server Conjecture

DGX agent

arXiv:2604.07240v1 Announce Type: cross Abstract: We introduce a code-based challenge for automated, open-ended mathematical discovery based on the k-server conjecture, a central open problem in com

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Learning Who Disagrees: Demographic Importance Weighting for Modeling Annotator Distributions with DiADEM

DGX agent

arXiv:2604.08425v1 Announce Type: cross Abstract: When humans label subjective content, they disagree, and that disagreement is not noise. It reflects genuine differences in perspective shaped by anno

safetyarxiv-cs-cl
10 Apr 2026
Safety

LINE: LLM-based Iterative Neuron Explanations for Vision Models

DGX agent

arXiv:2604.08039v1 Announce Type: new Abstract: Interpreting the concepts encoded by individual neurons in deep neural networks is a crucial step towards understanding their complex decision-making pr

safetyarxiv-cs-cv
10 Apr 2026
Safety

LLM-based Schema-Guided Extraction and Validation of Missing-Person Intelligence from Heterogeneous Data Sources

DGX agent

arXiv:2604.06571v1 Announce Type: cross Abstract: Missing-person and child-safety investigations rely on heterogeneous case documents, including structured forms, bulletin-style posters, and narrative

safetyarxiv-cs-ai
10 Apr 2026
Safety

MDP modeling for multi-stage stochastic programs

DGX agent

arXiv:2509.22981v2 Announce Type: replace Abstract: We study a class of multi-stage stochastic programs, which incorporate modeling features from Markov decision processes (MDPs). This class includes

safetyarxiv-cs-lg
10 Apr 2026
Model Releases

MedDialBench: Benchmarking LLM Diagnostic Robustness under Parametric Adversarial Patient Behaviors

DGX agent

arXiv:2604.06846v1 Announce Type: cross Abstract: Interactive medical dialogue benchmarks have shown that LLM diagnostic accuracy degrades significantly when interacting with non-cooperative patients,

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

MonoUNet: A Robust Tiny Neural Network for Automated Knee Cartilage Segmentation on Point-of-Care Ultrasound Devices

DGX agent

arXiv:2604.07780v1 Announce Type: cross Abstract: Objective: To develop a robust and compact deep learning model for automated knee cartilage segmentation on point-of-care ultrasound (POCUS) devices.

safetyarxiv-cs-cv
10 Apr 2026
Safety

Multi-Faceted Self-Consistent Preference Alignment for Query Rewriting in Conversational Search

DGX agent

arXiv:2604.06771v1 Announce Type: cross Abstract: Conversational Query Rewriting (CQR) aims to rewrite ambiguous queries to achieve more efficient conversational search. Early studies have predominant

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

Not All Tokens See Equally: Perception-Grounded Policy Optimization for Large Vision-Language Models

DGX agent

arXiv:2604.01840v2 Announce Type: replace Abstract: While Reinforcement Learning from Verifiable Rewards (RLVR) has advanced reasoning in Large Vision-Language Models (LVLMs), prevailing frameworks su

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks

DGX agent

arXiv:2604.08539v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as the de facto Reinforcement Learning (RL) objective driving recent advancements in Multimodal

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

PeReGrINE: Evaluating Personalized Review Fidelity with User Item Graph Context

DGX agent

arXiv:2604.07788v1 Announce Type: cross Abstract: We introduce PeReGrINE, a benchmark and evaluation framework for personalized review generation grounded in graph-structured user--item evidence. PeRe

model-releasesarxiv-cs-cl
10 Apr 2026
Research

Phantasia: Context-Adaptive Backdoors in Vision Language Models

DGX agent

arXiv:2604.08395v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have greatly enhanced the integration of visual perception and linguistic reasoning, driving rapid prog

researcharxiv-cs-cv
10 Apr 2026
Applications

PIArena: A Platform for Prompt Injection Evaluation

DGX agent

arXiv:2604.08499v1 Announce Type: cross Abstract: Prompt injection attacks pose serious security risks across a wide range of real-world applications. While receiving increasing attention, the communi

applicationsarxiv-cs-cl
10 Apr 2026
Safety

Plug-and-Play Logit Fusion for Heterogeneous Pathology Foundation Models

DGX agent

arXiv:2604.07779v1 Announce Type: new Abstract: Pathology foundation models (FMs) have become central to computational histopathology, offering strong transfer performance across a wide range of diagn

safetyarxiv-cs-cv
10 Apr 2026
Safety

ReflectRM: Boosting Generative Reward Models via Self-Reflection within a Unified Judgment Framework

DGX agent

arXiv:2604.07506v1 Announce Type: cross Abstract: Reward Models (RMs) are critical components in the Reinforcement Learning from Human Feedback (RLHF) pipeline, directly determining the alignment qual

safetyarxiv-cs-cl
10 Apr 2026
Safety

Reset-Free Reinforcement Learning for Real-World Agile Driving: An Empirical Study

DGX agent

arXiv:2604.07672v1 Announce Type: new Abstract: This paper presents an empirical study of reset-free reinforcement learning (RL) for real-world agile driving, in which a physical 1/10-scale vehicle le

safetyarxiv-cs-ro
10 Apr 2026
Hardware

RLBoost: Harvesting Preemptible Resources for Cost-Efficient Reinforcement Learning on LLMs

DGX agent

arXiv:2510.19225v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become essential for unlocking advanced reasoning capabilities in large language models (LLMs). RL workflows i

hardwarearxiv-cs-lg
10 Apr 2026
Safety

RoSHI: A Versatile Robot-oriented Suit for Human Data In-the-Wild

DGX agent

arXiv:2604.07331v1 Announce Type: cross Abstract: Scaling up robot learning will likely require human data containing rich and long-horizon interactions in the wild. Existing approaches for collecting

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

SE-Enhanced ViT and BiLSTM-Based Intrusion Detection for Secure IIoT and IoMT Environments

DGX agent

arXiv:2604.06254v1 Announce Type: cross Abstract: With the rapid growth of interconnected devices in Industrial and Medical Internet of Things (IIoT and MIoT) ecosystems, ensuring timely and accurate

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

Self-Debias: Self-correcting for Debiasing Large Language Models

DGX agent

arXiv:2604.08243v1 Announce Type: new Abstract: Although Large Language Models (LLMs) demonstrate remarkable reasoning capabilities, inherent social biases often cascade throughout the Chain-of-Though

safetyarxiv-cs-cl
10 Apr 2026
Safety

SeMoBridge: Semantic Modality Bridge for Efficient Few-Shot Adaptation of CLIP

DGX agent

arXiv:2509.26036v3 Announce Type: replace Abstract: While Contrastive Language-Image Pretraining (CLIP) excels at zero-shot tasks by aligning image and text embeddings, its performance in few-shot cla

safetyarxiv-cs-cv
10 Apr 2026
Safety

SIM1: Physics-Aligned Simulator as Zero-Shot Data Scaler in Deformable Worlds

DGX agent

arXiv:2604.08544v1 Announce Type: cross Abstract: Robotic manipulation with deformable objects represents a data-intensive regime in embodied learning, where shape, contact, and topology co-evolve in

safetyarxiv-cs-cv
10 Apr 2026
Tutorials

Steering the Verifiability of Multimodal AI Hallucinations

DGX agent

arXiv:2604.06714v1 Announce Type: new Abstract: AI applications driven by multimodal large language models (MLLMs) are prone to hallucinations and pose considerable risks to human users. Crucially, su

tutorialsarxiv-cs-ai
10 Apr 2026
Applications

Stop Listening to Me! How Multi-turn Conversations Can Degrade LLM Diagnostic Reasoning

DGX agent

arXiv:2603.11394v2 Announce Type: replace Abstract: Patients and clinicians are increasingly using chatbots powered by large language models (LLMs) for healthcare inquiries. While state-of-the-art LLM

applicationsarxiv-cs-cl
10 Apr 2026
Safety

SymptomWise: A Deterministic Reasoning Layer for Reliable and Efficient AI Systems

DGX agent

arXiv:2604.06375v1 Announce Type: new Abstract: AI-driven symptom analysis systems face persistent challenges in reliability, interpretability, and hallucination. End-to-end generative approaches ofte

safetyarxiv-cs-ai
10 Apr 2026
Safety

SYN-DIGITS: A Synthetic Control Framework for Calibrated Digital Twin Simulation

DGX agent

arXiv:2604.07513v1 Announce Type: cross Abstract: AI-based persona simulation -- often referred to as digital twin simulation -- is increasingly used for market research, recommender systems, and soci

safetyarxiv-cs-cl
10 Apr 2026
Model Releases

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

DGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

The Art of (Mis)alignment: How Fine-Tuning Methods Effectively Misalign and Realign LLMs in Post-Training

DGX agent

arXiv:2604.07754v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) raises significant ethical and safety concerns. While LLM alignment techniques are adopted to improve m

safetyarxiv-cs-cl
10 Apr 2026
Safety

The Sustainability Gap in Robotics: A Large-Scale Survey of Sustainability Awareness in 50,000 Research Articles

DGX agent

arXiv:2604.07921v1 Announce Type: new Abstract: We present a large-scale survey of sustainability communication and motivation in robotics research. Our analysis covers nearly 50,000 open-access paper

safetyarxiv-cs-ro
10 Apr 2026
Safety

Towards the Development of an LLM-Based Methodology for Automated Security Profiling in Compliance with Ukrainian Cybersecurity Regulations

DGX agent

arXiv:2604.06274v1 Announce Type: cross Abstract: In recent years, the pace of development of information technology in various areas has increased drastically, forcing cybersecurity specialists to co

safetyarxiv-cs-ai
10 Apr 2026
Model Releases

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

DGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

WASD: Locating Critical Neurons as Sufficient Conditions for Explaining and Controlling LLM Behavior

DGX agent

arXiv:2603.18474v2 Announce Type: replace Abstract: Precise behavioral control of large language models (LLMs) is critical for complex applications. However, existing methods often incur high training

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

What Drives Representation Steering? A Mechanistic Case Study on Steering Refusal

DGX agent

arXiv:2604.08524v1 Announce Type: cross Abstract: Applying steering vectors to large language models (LLMs) is an efficient and effective model alignment technique, but we lack an interpretable explan

safetyarxiv-cs-cl
10 Apr 2026
Applications

Dynamics Models for Offline Hyperparameter Selection in Real-World RL

DGX agent

arXiv:2608.11349v1 Announce Type: cross Abstract: A key obstacle to deploying reinforcement learning in real-world systems is hyperparameter selection, particularly when simulators are unavailable and

applicationsarxiv-cs-ai
13 Aug 2026
Safety

Enhancing Visual Domain Robustness in Behaviour Cloning via Saliency-Guided Augmentation

DGX agent

arXiv:2608.11870v1 Announce Type: new Abstract: In vision-based behavior cloning (BC), conventional image augmentations such as Random Crop and Color Jitter often fall short under substantial visual d

safetyarxiv-cs-ro
13 Aug 2026
← Previous
1…201202203204205…233
Next →