AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Geometry of Reason: Spectral Signatures of Valid Mathematical Reasoning

DGX agent

arXiv:2601.00791v2 Announce Type: replace-cross Abstract: Verifying whether a language model is genuinely reasoning or pattern-matching remains an open problem: learned verifiers are expensive, and ou

researcharxiv-cs-ai
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Goal-Autopilot: A Verifiable Anti-Fabrication Firewall for Unattended Long-Horizon Agents

DGX agent

arXiv:2606.11688v1 Announce Type: cross Abstract: Long-horizon LLM agents are not trusted to run unattended: with no human watching, they confidently report success they never verified. We treat hones

agentsarxiv-cs-ai
11 Jun 2026
Safety

GPO: Learning from Critical Steps to Improve LLM Reasoning

DGX agent

arXiv:2509.16456v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used in various domains, showing impressive potential on different tasks. Recently, reasoning LLMs hav

safetyarxiv-cs-ai
11 Jun 2026
Safety

Grammar-Constrained Decoding Can Jailbreak LLMs into Generating Malicious Code

DGX agent

arXiv:2606.11817v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code generation, raising concerns that they may be misused to produce malicious code. Meanwhile

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Grounding Computer Use Agents on Human Demonstrations

DGX agent

arXiv:2511.07332v2 Announce Type: replace-cross Abstract: Building reliable computer-use agents requires grounding: accurately connecting natural language instructions to the correct on-screen element

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Harness In-Context Operator Learning with Chain of Operators

DGX agent

arXiv:2606.12318v1 Announce Type: cross Abstract: Neural operators approximate mappings between function spaces, but often generalize poorly to other operators and usually require fine-tuning or retra

researcharxiv-cs-ai
11 Jun 2026
Safety

HERO: Hindsight-Enhanced Reflection from Environment Observations for Agentic Self-Distillation

DGX agent

arXiv:2606.11559v1 Announce Type: new Abstract: Reinforcement learning typically improves multi-turn agent capabilities through the terminal outcome of the trajectories, which makes it difficult to de

safetyarxiv-cs-ai
11 Jun 2026
Safety

Hey Chat, Can You Teach Me? Structuring Socratic Dialogue for Human Learning in the Wild

DGX agent

arXiv:2606.11744v1 Announce Type: cross Abstract: Large language models are now widely used for everyday learning, but the underlying interactions are typically unstructured chats rather than followin

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Hubs or Fringes: Pretraining Data Selection via Web Graph Centrality

DGX agent

arXiv:2606.11499v1 Announce Type: cross Abstract: The performance of modern language models depends critically on pretraining data composition. Yet existing data selection methods rely on auxiliary cl

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

Human-Enhanced Loop Modeling (HELM): Agent-Based Finite Element Modeling of Concrete Bridge Barriers

DGX agent

arXiv:2606.12025v1 Announce Type: new Abstract: Finite element (FE) modeling of safety-critical infrastructure such as bridge barriers requires high-fidelity nonlinear dynamic analysis, yet the curren

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Human-Guided Agentic AI for Multimodal Clinical Prediction: Lessons from the AgentDS Healthcare Benchmark

DGX agent

arXiv:2602.19502v2 Announce Type: replace Abstract: Agentic AI systems are increasingly capable of autonomous data science workflows, yet clinical prediction tasks demand domain expertise that purely

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

ICA Lens: Interpreting Language Models Without Training Another Dictionary

DGX agent

arXiv:2606.11722v1 Announce Type: cross Abstract: Finding interpretable directions in language-model representations is critical for understanding and controlling model behavior. Sparse autoencoders (

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

Illumination-Robust Camera-Based Heart-Rate Estimation for Physiological Sensing in Robots

DGX agent

arXiv:2606.12378v1 Announce Type: cross Abstract: Physiological awareness is important for service, social, and assistive robots that interact with humans in everyday environments. Remote photoplethys

safetyarxiv-cs-ai
11 Jun 2026
Safety

Implicit Neural Representations of Individual Behavior

DGX agent

arXiv:2606.12200v1 Announce Type: cross Abstract: We study policy representation learning from unlabeled multi-policy behavioral data. Each episode is generated by a fixed policy, but policy labels ar

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Improving Detection of Rare Nodes in Hierarchical Multi-Label Learning

DGX agent

arXiv:2602.08986v2 Announce Type: replace-cross Abstract: In hierarchical multi-label classification, a persistent challenge is enabling model predictions to reach deeper levels of the hierarchy for m

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

Improving Generalization and Data Efficiency with Diffusion in Offline Multi-agent RL

DGX agent

arXiv:2307.01472v2 Announce Type: replace Abstract: We present a novel Diffusion Offline Multi-agent Model (DOM2) for offline Multi-Agent Reinforcement Learning (MARL). Different from existing algorit

safetyarxiv-cs-ai
11 Jun 2026
Tutorials

Information-Theoretic Decomposition for Multimodal Interaction Learning

DGX agent

arXiv:2606.11614v1 Announce Type: cross Abstract: Multimodal learning hinges on capturing redundant, unique, and synergistic information across modalities, which collectively constitute multimodal int

tutorialsarxiv-cs-ai
11 Jun 2026
Hardware

INFRAMIND: Infrastructure-Aware Multi-Agent Orchestration

DGX agent

arXiv:2606.11440v1 Announce Type: new Abstract: Existing multi-agent LLM orchestration methods, ranging from brute-force ensembles to learned routers, select models and topologies based on task and mo

hardwarearxiv-cs-ai
11 Jun 2026
Safety

IntElicit: Eliciting and Assessing Contextualized Creativity via Dialogue Policy Optimization

DGX agent

arXiv:2606.12086v1 Announce Type: new Abstract: Contextualized assessment offers high ecological validity for evaluating creativity but introduces a critical challenge: observed performance may be con

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Intelligent Automation for Embodied Benchmark Construction: Pipelines, Embodiments, Simulators, and Trends

DGX agent

arXiv:2606.12207v1 Announce Type: cross Abstract: Embodied intelligence now spans navigation, household assistance, manipulation, autonomous driving, aerial agents, and multimodal large-model control.

model-releasesarxiv-cs-ai
11 Jun 2026
Applications

Internet of Everything in the 6G Era: Paradigms, Enablers, Potentials and Future Directions

DGX agent

arXiv:2604.25018v2 Announce Type: replace-cross Abstract: The Internet of Everything (IoE) represents an evolution of the Internet of Things (IoT) by integrating people, data, processes, and things in

applicationsarxiv-cs-ai
11 Jun 2026
Research

Irresponsible AI: big tech's influence on AI research and associated impacts

DGX agent

arXiv:2512.03077v2 Announce Type: replace-cross Abstract: The accelerated development, deployment and adoption of artificial intelligence systems has been fuelled by the increasing presence of big tec

researcharxiv-cs-ai
11 Jun 2026
Agents

ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories

DGX agent

arXiv:2606.11520v1 Announce Type: cross Abstract: Training capable OS agents requires data that simultaneously captures structured user intents, multi-turn task delegation, and grounded tool execution

agentsarxiv-cs-ai
11 Jun 2026
Safety

JailbreakOPT: Tool-Assisted Iterative Jailbreak Prompt Optimization

DGX agent

arXiv:2606.11425v1 Announce Type: cross Abstract: Jailbreak attacks expose persistent safety weaknesses in large language models (LLMs), but existing stateless single-turn methods face a trade-off: ha

safetyarxiv-cs-ai
11 Jun 2026
Agents

Knowing When to Ask: Self-Gated Clarification for Hierarchical Language Agents

DGX agent

arXiv:2606.11349v1 Announce Type: new Abstract: In hierarchical reasoning, failures often originate at intermediate decision points where the agent commits to a wrong branch without recognizing that i

agentsarxiv-cs-ai
11 Jun 2026
Applications

LaQual: An Automated Framework for LLM App Quality Evaluation

DGX agent

arXiv:2508.18636v2 Announce Type: replace-cross Abstract: Representing a new paradigm in software distribution, LLM app stores are rapidly emerging, offering users diverse choices for content generati

applicationsarxiv-cs-ai
11 Jun 2026
Local Ai

LASA: A Weak Supervision Method for Open-Vocabulary Scene Sketch Semantic Segmentation

DGX agent

arXiv:2606.11837v1 Announce Type: cross Abstract: Open-vocabulary scene sketch semantic segmentation aims to assign dense semantic labels to sparse line drawings based on flexible category vocabularie

local-aiarxiv-cs-ai
11 Jun 2026
Safety

Latent World Recovery for Multimodal Learning with Missing Modalities

DGX agent

arXiv:2606.12362v1 Announce Type: cross Abstract: We study multimodal learning under missing modalities, with particular motivation from bioscience applications in which heterogeneous modalities are o

safetyarxiv-cs-ai
11 Jun 2026
Local Ai

Layer-Isolated Evaluation: Gating the Deterministic Scaffold of a Production LLM Agent with a No-LLM, Regression-Locked Test Harness

DGX agent

arXiv:2606.11686v1 Announce Type: cross Abstract: End-to-end task-success is the dominant way to evaluate LLM agents, but one aggregate number tells you that an agent regressed, not where. We present

local-aiarxiv-cs-ai
11 Jun 2026
Safety

Learning to Inject: Automated Prompt Injection via Reinforcement Learning

DGX agent

arXiv:2602.05746v2 Announce Type: replace-cross Abstract: Prompt injection is a critical vulnerability in LLM agents, yet the strongest methods still rely on human red-teamers and hand-crafted prompts

safetyarxiv-cs-ai
11 Jun 2026
Hardware

Litespark Inference For CPUs: Ultra-Fast SIMD Framework for Ternary (1.58-bit) Language Models

DGX agent

arXiv:2605.06485v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have transformed artificial intelligence, but their computational requirements remain prohibitive for most users.

hardwarearxiv-cs-ai
11 Jun 2026
Agents

LLMs+Graphs: Toward Graph-Native, Synergistic AI Systems

DGX agent

arXiv:2606.11560v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced rapidly, but their limitations in structured and multi-hop reasoning underscore the need for graph-native,

agentsarxiv-cs-ai
11 Jun 2026
Research

LSTM-Based Detection of Structural Breaks in Property Insurance Loss Reserving: A Climate-Informed Approach

DGX agent

arXiv:2606.11463v1 Announce Type: cross Abstract: Accurate loss reserving is foundational to insurer solvency, yet accelerating climate driven catastrophes systematically violate the stability assumpt

researcharxiv-cs-ai
11 Jun 2026
Research

LSTM based IoT Device Identification

DGX agent

arXiv:2304.13905v2 Announce Type: replace-cross Abstract: While the use of the Internet of Things is becoming more and more popular, many security vulnerabilities are emerging with the large number of

researcharxiv-cs-ai
11 Jun 2026
Safety

LUCID: Learning Embodiment-Agnostic Intent Models from Unstructured Human Videos for Scalable Dexterous Robot Skill Acquisition

DGX agent

arXiv:2606.11628v1 Announce Type: cross Abstract: The most widely-adopted robot learning pipelines today learn skills from robot demonstrations or structured human data, which are expensive to collect

safetyarxiv-cs-ai
11 Jun 2026
Research

Lung-R1: A Knowledge Graph-Guided LLM for Pulmonary Diagnostic Reasoning

DGX agent

arXiv:2606.11675v1 Announce Type: new Abstract: Diagnosing pulmonary diseases requires integrating heterogeneous evidence amid phenotypic variability and cross-disease overlap. Although large language

researcharxiv-cs-ai
11 Jun 2026
Model Releases

Lung-SRAD: Spectral-Aware Regularized Audio DASS with Dual-Axis Patch-Mix Contrastive Learning for Respiratory Sound Classification

DGX agent

arXiv:2606.11922v1 Announce Type: cross Abstract: Recent respiratory sound classification (RSC) studies largely rely on CLS-token driven self-attention architectures such as the Audio Spectrogram Tran

model-releasesarxiv-cs-ai
11 Jun 2026
Research

MA-DLE: Speech-based Automatic Depression Level Estimation via Memory Augmentation

DGX agent

arXiv:2606.11197v1 Announce Type: cross Abstract: Speech-based automatic estimation of depression levels is essential for enabling early detection and timely intervention, particularly in resource-con

researcharxiv-cs-ai
11 Jun 2026
Local Ai

Making Foresight Actionable: Repurposing Representation Alignment in World Action Models

DGX agent

arXiv:2606.12217v1 Announce Type: cross Abstract: World Action Models (WAMs) offer a promising route for robot manipulation by using video generation models to model future scene evolution before prod

local-aiarxiv-cs-ai
11 Jun 2026
Model Releases

Making Models Unmergeable via Scaling-Sensitive Loss Landscape

DGX agent

arXiv:2601.21898v2 Announce Type: replace Abstract: The rise of model hubs has made it easier to access reusable model components, making model merging a practical tool for combining capabilities. Yet

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Mapping Scientific Literature with Large Language Models and Topic Modeling

DGX agent

arXiv:2510.16152v2 Announce Type: replace-cross Abstract: Scientific literature is increasingly fragmented by disciplinary boundaries, specialized terminology, and potentially sparse keyword systems,

researcharxiv-cs-ai
11 Jun 2026
Model Releases

MARIC: Multi-Agent Reasoning for Image Classification

DGX agent

arXiv:2509.14860v2 Announce Type: replace-cross Abstract: Image classification has traditionally relied on parameter-intensive model training, requiring large-scale annotated datasets and extensive fi

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Market Design for AI: Beyond the Copyright Binary

DGX agent

arXiv:2606.12260v1 Announce Type: cross Abstract: How can we design a market of human-generated content for use in training AI models that both enables technological progress and preserves individual

researcharxiv-cs-ai
11 Jun 2026
Research

Mathematical perspective on genetic algorithms with optimization guided operators

DGX agent

arXiv:2606.12279v1 Announce Type: cross Abstract: Recent work in ML applies genetic algorithms at inference time to iteratively improve solutions to optimization problems. The basic mutation and recom

researcharxiv-cs-ai
11 Jun 2026
Model Releases

MedCTA: A Benchmark for Clinical Tool Agents

DGX agent

arXiv:2606.11702v1 Announce Type: cross Abstract: To make clinically grounded decisions, medical AI agents are expected to go beyond simple recognition and be capable of tool retrieval, evidence acqui

model-releasesarxiv-cs-ai
11 Jun 2026
Research

MentisOculi: Revealing the Limits of Reasoning with Mental Imagery

DGX agent

arXiv:2602.02465v2 Announce Type: replace Abstract: Frontier models are transitioning from multimodal large language models (MLLMs) that merely ingest visual information to unified multimodal models (

researcharxiv-cs-ai
11 Jun 2026
Model Releases

Metadata-Aware Multi-Prompt Reasoning for Zero-Shot Accident Understanding

DGX agent

arXiv:2606.12047v1 Announce Type: cross Abstract: In this paper, we address the problem of zero-shot understanding of accidents from surveillance videos by identifying when an impact event occurs, wha

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Mind the Perspective: Let's Reason Recursively for Theory of Mind

DGX agent

arXiv:2606.11724v1 Announce Type: new Abstract: Theory of Mind (ToM) reasoning requires inferring agents' beliefs from partial and asymmetric observations, which remains an open challenge for LLMs. Ex

model-releasesarxiv-cs-ai
11 Jun 2026
← Previous
1…164165166167168…448
Next →