AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
22,134 results
Safety

Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains

DGX agent

arXiv:2605.19940v1 Announce Type: new Abstract: Foundation models are increasingly deployed in socially sensitive domains such as education, mental health, and caregiving, where failures are often cum

safetyarxiv-cs-ai
20 May 2026
Tutorials
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Smartphone-based Circular Plot Sampling for Forest Inventory

DGX agent

arXiv:2605.19213v1 Announce Type: new Abstract: Circular sample plots are a cornerstone of forest inventory, yet accurate measurement of tree diameter at breast height (DBH) and spatial location withi

tutorialsarxiv-cs-cv
20 May 2026
Model Releases

Sonar-TS: Search-Then-Verify Natural Language Querying for Time Series Databases

DGX agent

arXiv:2602.17001v2 Announce Type: replace Abstract: Natural Language Querying for Time Series Databases (NLQ4TSDB) aims to assist non-expert users retrieve meaningful events, intervals, and summaries

model-releasesarxiv-cs-ai
20 May 2026
Agents

Stop Drawing Scientific Claims from LLM Social Simulations Without Robustness Audits

DGX agent

arXiv:2605.18890v1 Announce Type: cross Abstract: The scientific claims drawn from LLM social simulations should be no stronger than the robustness audits that support them. Generative agents bring ne

agentsarxiv-cs-ai
20 May 2026
Model Releases

Synthetic Data Generation for Brain-Computer Interfaces: Overview, Benchmarking, and Future Directions

DGX agent

arXiv:2603.12296v2 Announce Type: replace-cross Abstract: Deep learning has achieved transformative performance across diverse domains, largely driven by large-scale and high-quality training data. In

model-releasesarxiv-cs-ai
20 May 2026
Safety

The Accessibility Capability Boundary: Operational Limits and Expansion Potential of AI-Generated Browser-Native Accessibility Systems

DGX agent

arXiv:2605.19638v1 Announce Type: cross Abstract: As large language models (LLMs) demonstrate increasing competence in synthesizing functional user interfaces, a fundamental question emerges in access

safetyarxiv-cs-ai
20 May 2026
Safety

Toward an AI-Powered Computational Testbed for Workforce Policy

DGX agent

arXiv:2605.19064v1 Announce Type: cross Abstract: Workforce transformations are difficult to forecast and costly to mismanage. In particular, the integration of artificial intelligence into knowledge

safetyarxiv-cs-ai
20 May 2026
Model Releases

Trajectory Planning and Control near the Limits: an Open Experimental Benchmark on the RoboRacer Platform

DGX agent

arXiv:2605.19881v1 Announce Type: new Abstract: We present a modular framework to benchmark new and existing methods for trajectory planning and control in high-acceleration maneuvers that push autono

model-releasesarxiv-cs-ro
20 May 2026
Agents

When Web Apps Heal Themselves: A MAPE-K Based Approach to Fault Tolerance and Adaptive Recovery

DGX agent

arXiv:2605.19261v1 Announce Type: cross Abstract: Ensuring the reliability and resilience of modern web applications remains a critical challenge due to increasing system complexity and dynamic runtim

agentsarxiv-cs-ai
20 May 2026
Model Releases

ZeroSearch: Incentivize the Search Capability of LLMs without Searching

DGX agent

arXiv:2505.04588v3 Announce Type: replace Abstract: Effective information searching is essential for enhancing the reasoning and generation capabilities of large language models (LLMs). Recent researc

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

4DLidarOpen: An Open 4D FMCW Lidar Dataset for Motion-Aware Autonomous Driving

DGX agent

arXiv:2605.18074v1 Announce Type: new Abstract: We present 4DLidarOpen, a large-scale open multi-modal dataset for autonomous driving, centered on 4D frequency-modulated continuous-wave (FMCW) Lidar s

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

A Feature-Driven Framework for Software Fault Prediction

DGX agent

arXiv:2605.17611v1 Announce Type: cross Abstract: Software fault prediction (SFP) is a critical task in software engineering, enabling early identification of faults in modules to improve software qua

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

A Pilot Benchmark for NL-to-FOL Translation in Planetary Exploration

DGX agent

arXiv:2605.17911v1 Announce Type: new Abstract: Future planetary exploration envisions autonomous robotic agents operating under severe communication constraints, without global positioning, and with

model-releasesarxiv-cs-cl
19 May 2026
Applications

Active Budget Allocation for Efficient Scaling Law Estimation via Surrogate-Guided Pruning

DGX agent

arXiv:2605.17234v1 Announce Type: new Abstract: Predicting model performance at larger scales enables the design of training strategies and architectures tailored to specific performance targets. Empi

applicationsarxiv-cs-lg
19 May 2026
Agents

AgentArk: Distilling Multi-Agent Intelligence into a Single LLM Agent

DGX agent

arXiv:2602.03955v2 Announce Type: replace Abstract: While large language model (LLM) multi-agent systems achieve superior reasoning performance through iterative debate, practical deployment is limite

agentsarxiv-cs-ai
19 May 2026
Safety

AI Agents May Always Fall for Prompt Injections

DGX agent

arXiv:2605.17634v1 Announce Type: cross Abstract: Prompt injection is the most critical vulnerability in deployed AI agents. Despite recent progress, we show that the prevailing defense paradigm (data

safetyarxiv-cs-cl
19 May 2026
Safety

Alignment Drift in Long-Term Human-LLM Interaction: A Mechanism-Oriented Framework

DGX agent

arXiv:2605.16516v1 Announce Type: cross Abstract: Long-term interaction with LLM-based systems may produce alignment drift: a gradual process in which system outputs become less constrained by the use

safetyarxiv-cs-ai
19 May 2026
Safety

ARROW: Augmented Replay for RObust World models

DGX agent

arXiv:2603.11395v2 Announce Type: replace-cross Abstract: Continual reinforcement learning challenges agents to acquire new skills while retaining previously learned ones with the goal of improving pe

safetyarxiv-cs-ai
19 May 2026
Model Releases

Babel: Jailbreaking Safety Attention via Obfuscation Distribution Optimized Sampling

DGX agent

arXiv:2605.17971v1 Announce Type: cross Abstract: Despite rigorous safety alignment, Large Language Models (LLMs) remain vulnerable to jailbreak attacks. Existing black-box methods often rely on heuri

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Bench2Drive-Robust: Benchmarking Closed-Loop Autonomous Driving under Deployment Perturbations

DGX agent

arXiv:2605.18059v1 Announce Type: new Abstract: Robustness is a critical requirement for deploying autonomous driving systems in the real world. Existing robustness benchmarks for autonomous driving h

model-releasesarxiv-cs-ro
19 May 2026
Safety

Beyond Compliance: How AI Could Help Creative Writers by Refusing Them

DGX agent

arXiv:2605.16272v1 Announce Type: cross Abstract: Mainstream creativity support design prioritizes compliant AI for seamless writing interactions, but concerns over inappropriate AI reliance highlight

safetyarxiv-cs-ai
19 May 2026
Model Releases

Breaking Annotation Barriers: Generalized Video Quality Assessment via Ranking-based Self-Supervision

DGX agent

arXiv:2505.03631v4 Announce Type: replace Abstract: Video quality assessment (VQA) is essential for quantifying perceptual quality in various video processing workflows, spanning from camera capture s

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Can LLMs Generate and Solve Linguistic Olympiad Puzzles?

DGX agent

arXiv:2509.21820v2 Announce Type: replace Abstract: In this paper, we introduce a combination of novel and exciting tasks: the solution and generation of linguistic puzzles. We focus on puzzles used i

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

CAREBench: Evaluating LLMs' Emotion Understanding by Assessing Cognitive Appraisal Reasoning

DGX agent

arXiv:2605.17176v1 Announce Type: new Abstract: Emotion understanding is a core capability for LLMs to interact effectively with humans, yet existing evaluation paradigms rely on discrete emotion labe

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

CayleyPy RL: Pathfinding and Reinforcement Learning on Cayley Graphs

DGX agent

arXiv:2502.18663v3 Announce Type: replace Abstract: This paper is the second in a series of studies on developing efficient artificial intelligence-based approaches to pathfinding on extremely large g

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Continual Learning for VLMs: A Survey and Taxonomy Beyond Forgetting

DGX agent

arXiv:2508.04227v2 Announce Type: replace Abstract: Vision-language models (VLMs) and the recent surge of Multimodal Large Language Models (MLLMs) have revolutionized artificial intelligence with unpr

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

CooT: Learning to Coordinate In-Context with Coordination Transformers

DGX agent

arXiv:2506.23549v3 Announce Type: replace Abstract: Effective coordination among unfamiliar partners remains a major challenge in multi-agent systems. Existing approaches, such as population-based met

model-releasesarxiv-cs-ai
19 May 2026
Safety

Deep sequence models tend to memorize geometrically; it is unclear why

DGX agent

arXiv:2510.26745v3 Announce Type: replace-cross Abstract: Deep sequence models are said to store atomic facts predominantly in the form of associative memory: a brute-force lookup of co-occurring enti

safetyarxiv-cs-ai
19 May 2026
Model Releases

EgoIntrospect: An Egocentric Dataset and Benchmark for User-Centric Internal State Reasoning

DGX agent

arXiv:2605.17262v1 Announce Type: new Abstract: Despite extensive efforts on egocentric video datasets and benchmarks, understanding users' internal states, which is crucial for enabling seamless AI a

model-releasesarxiv-cs-cv
19 May 2026
Safety

Estimating Item Difficulty with Large Language Models as Experts

DGX agent

arXiv:2605.18562v1 Announce Type: cross Abstract: Accurate estimates of item difficulty are essential for valid assessment and effective adaptive learning. However, for newly created tasks, response d

safetyarxiv-cs-ai
19 May 2026
Model Releases

Evaluating Cognitive Age Alignment in Interactive AI Agents

DGX agent

arXiv:2605.17894v1 Announce Type: new Abstract: While agentic AI and its core multimodal large language models (MLLMs) have demonstrated remarkable promise in language and visual reasoning across doma

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Evolve the Method, Not the Prompts: Evolutionary Synthesis of Jailbreak Attacks on LLMs

DGX agent

arXiv:2511.12710v2 Announce Type: replace Abstract: Automated red teaming frameworks for Large Language Models (LLMs) have become increasingly sophisticated, yet many still formulate attack optimizati

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

EvoMemBench: Benchmarking Agent Memory from a Self-Evolving Perspective

DGX agent

arXiv:2605.18421v1 Announce Type: cross Abstract: Recent benchmarks for Large Language Model (LLM) agents mainly evaluate reasoning, planning, and execution. However, memory is also essential for agen

model-releasesarxiv-cs-ai
19 May 2026
Safety

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes

DGX agent

arXiv:2605.16303v1 Announce Type: cross Abstract: Large language models (LLM) agents may offer tools to predict human responses to surveys. A common technique for defining these agents uses only demog

safetyarxiv-cs-ai
19 May 2026
Agents

Generative AI Advertising as a Problem of Trustworthy Commercial Intervention

DGX agent

arXiv:2605.18673v1 Announce Type: cross Abstract: Major deployed generative AI advertising systems preserve a visible boundary between commercial content and AI-generated responses. Yet empirical rese

agentsarxiv-cs-cl
19 May 2026
Model Releases

Generative Artificial Intelligence for Literature Reviews

DGX agent

arXiv:2605.16475v1 Announce Type: cross Abstract: Generative artificial intelligence (GenAI), based on large-language models (LLMs), such as ChatGPT, has taken organizations, academia, and the public

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Graph Embedding in the Graph Fractional Fourier Transform Domain

DGX agent

arXiv:2508.02383v2 Announce Type: replace Abstract: Spectral graph embedding plays a critical role in graph representation learning by generating low-dimensional vector representations from graph spec

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

GVGAI-LLM: Evaluating Large Language Model Agents with Infinite Games

DGX agent

arXiv:2508.08501v3 Announce Type: replace Abstract: We introduce GVGAI-LLM, a video game benchmark for evaluating the reasoning and problem-solving capabilities of large language models (LLMs). Built

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

HalluScore: Large Language Model Hallucination Question Answering Benchmark

DGX agent

arXiv:2605.17007v1 Announce Type: new Abstract: Large language models (LLMs) have achieved remarkable progress in natural language generation, but remain susceptible to hallucination. In response to g

model-releasesarxiv-cs-cl
19 May 2026
Tutorials

High-Resolution Reference Image Assisted Volumetric Super-Resolution of Cardiac Diffusion Weighted Imaging

DGX agent

arXiv:2310.20389v2 Announce Type: replace-cross Abstract: Diffusion Tensor Cardiac Magnetic Resonance (DT-CMR) is the only in vivo method to non-invasively examine the microstructure of the human hear

tutorialsarxiv-cs-cv
19 May 2026
Model Releases

HTSC-2025: A Benchmark Dataset of Ambient-Pressure High-Temperature Superconductors for AI-Driven Critical Temperature Prediction

DGX agent

arXiv:2506.03837v2 Announce Type: replace-cross Abstract: The discovery of high-temperature superconducting materials holds great significance for human industry and daily life. In recent years, resea

model-releasesarxiv-cs-ai
19 May 2026
Tutorials

Inventorship in AI-Assisted Inventions: Designing an Experiment to Shape Case Law

DGX agent

arXiv:2605.16528v1 Announce Type: cross Abstract: The latest improvements in artificial intelligence (AI) raise new challenges for intellectual property laws, particularly concerning the inventorship

tutorialsarxiv-cs-ai
19 May 2026
Tutorials

iPOE: Interpretable Prompt Optimization via Explanations

DGX agent

arXiv:2605.18113v1 Announce Type: new Abstract: Prompt optimization has often been framed as a discrete search problem to find high-performing and robust instructions for an LLM. However, the search r

tutorialsarxiv-cs-cl
19 May 2026
Model Releases

KVCapsule: Efficient Sequential KV Cache Compression for Vision-Language Models with Asymmetric Redundancy

DGX agent

arXiv:2605.16439v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have emerged as a critical and fast-growing extension of Large Language Models (LLMs) that enable multimodal reasoning t

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Latency-Aware Deep Learning Benchmark for Real-Time Cyber-Physical Attack and Fault Classification in Inverter-Dominated Power Grids

DGX agent

arXiv:2605.17256v1 Announce Type: cross Abstract: This work introduces a latency-aware benchmarking framework for evaluating deep learning models in power system anomaly detection using high-fidelity,

model-releasesarxiv-cs-ai
19 May 2026
Tutorials

Learning What Evaluators Value: A Reliable Approach to Modeling Evaluator Preferences

DGX agent

arXiv:2605.16615v1 Announce Type: new Abstract: In many applications, human and LLM evaluators use assessments of relevant criteria to create an overall evaluation for an item or individual. For examp

tutorialsarxiv-cs-lg
19 May 2026
Safety

Linguistic Uncertainty and Reply Engagement on X: A Cross-Domain Replication of the Uncertainty-Reply Asymmetry

DGX agent

arXiv:2605.16289v1 Announce Type: cross Abstract: Linguistic uncertainty is common in social media, but its relationship with engagement remains unclear across languages and topics. Using 2,258 Englis

safetyarxiv-cs-cl
19 May 2026
Model Releases

LiTS: A Modular Framework for LLM Tree Search

DGX agent

arXiv:2603.00631v2 Announce Type: replace Abstract: LiTS is a modular Python framework for LLM reasoning via tree search. It decomposes tree search into three reusable components (Policy, Transition,

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…437438439440441…462
Next →