AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Model Releases

What Does ODRL Mean? A Cross-Level Ontological Grounding of Permissions, Prohibitions, and Duties in UFO-L

DGX agent

arXiv:2606.24344v1 Announce Type: cross Abstract: ODRL policy evaluators produce verdicts, but say nothing about the normative positions a policy brings into existence, the authority structures those

model-releasesarxiv-cs-ai
24 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

When AI Meets Finance (StockAgent): Large Language Model-based Stock Trading in Simulated Real-world Environments

DGX agent

arXiv:2407.18957v5 Announce Type: replace-cross Abstract: Can AI Agents simulate real-world trading environments to investigate the impact of external factors on stock trading activities (e.g., macroe

safetyarxiv-cs-ai
24 Jun 2026
Safety

When CQs Go Wrong: Challenges in CQ Verification with OE-Assist

DGX agent

arXiv:2606.24619v1 Announce Type: new Abstract: Competency Questions (CQs) are the central component of CQ-verification, an established process in which an ontology is evaluated against a set of natur

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

When Helpfulness Overrides Causal Caution: Context-Dependent Suppression and Recovery in LLMs

DGX agent

arXiv:2606.24370v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into decision-support roles in business and policy contexts. While prior benchmark studies have

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

When Preferences Fail to Become Incentives: A Utility-Behavior Gap in Large Language Models

DGX agent

arXiv:2606.22974v2 Announce Type: replace Abstract: Recent work on preference elicitation in large language models (LLMs) has demonstrated that, when given a series of choices between two outcomes, LL

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

When Retrieval Metrics Mislead: Measuring Policy Signal in Long-Horizon Tool-Use Agents

DGX agent

arXiv:2606.23937v1 Announce Type: cross Abstract: Exact-match retrieval recall is often used as a proxy for whether a retriever supplies useful policy context to a downstream decision model. We test t

model-releasesarxiv-cs-ai
24 Jun 2026
Local Ai

World Models in Pieces: Structural Certification for General Agents

DGX agent

arXiv:2606.24842v1 Announce Type: new Abstract: In the big-world regime, agents cannot be universally capable and their ability is inevitably specialized across a world model in pieces. Consequently,

local-aiarxiv-cs-ai
24 Jun 2026
Research

Zero-Shot Test-Time Canonicalization using Out-of-Distribution Scoring

DGX agent

arXiv:2606.24178v1 Announce Type: cross Abstract: Pretrained vision models often misclassify inputs that are rotated, scaled, or sheared, even though these affine transformations leave the object clas

researcharxiv-cs-ai
24 Jun 2026
Model Releases

ZONOS2 Technical Report

DGX agent

arXiv:2606.24320v1 Announce Type: cross Abstract: We present ZONOS2 8B, our latest TTS model, which achieves state-of-the-art naturalness, prosody, and voice cloning fidelity. We improve upon Zonos-v0

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents

DGX agent

arXiv:2606.12320v1 Announce Type: new Abstract: Enterprise security was built to govern data boundaries: the protected surface was data at rest and in transit, and the controls -- access control, data

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

A Lightweight Multi-Agent Framework for Automated Concrete Barrier Design

DGX agent

arXiv:2606.12040v1 Announce Type: new Abstract: The design of reinforced concrete highway barriers is a safety-critical process that requires strict compliance with regulatory provisions such as the A

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

A Physics-Inspired Optimizer: Velocity Regularized Adam

DGX agent

arXiv:2505.13196v3 Announce Type: replace-cross Abstract: We introduce Velocity-Regularized Adam (VRAdam), a physics-inspired optimizer for training deep neural networks that draws on ideas from quart

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

A Survey of Reasoning and Agentic Systems in Time Series with Large Language Models

DGX agent

arXiv:2509.11575v3 Announce Type: replace Abstract: Time series reasoning treats time as a first-class axis and incorporates intermediate evidence directly into the answer. This survey defines the pro

safetyarxiv-cs-ai
11 Jun 2026
Applications

A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data

DGX agent

arXiv:2601.17717v3 Announce Type: replace Abstract: Large Language Models (LLMs) have emerged as powerful tools for generating data across various modalities. By transforming data from a scarce resour

applicationsarxiv-cs-ai
11 Jun 2026
Model Releases

Adapting Prithvi-EO for Fallow Detection for Food-Water Nexus: ViT-Adapter Necks and Parameter-Efficient Backbone tuning of Geospatial Foundation Model

DGX agent

arXiv:2606.12218v1 Announce Type: cross Abstract: Understanding spatial distribution of fallow land is important for optimizing the food-water (FW) nexus, given fallowing's role in crop rotation and w

model-releasesarxiv-cs-ai
11 Jun 2026
Research

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents

DGX agent

arXiv:2606.11219v1 Announce Type: cross Abstract: Audio language models (ALMs) are increasingly used for speech-based understanding, yet their ability to perform semantic reasoning beyond transcriptio

researcharxiv-cs-ai
11 Jun 2026
Agents

Agentic Environment Engineering for Large Language Models: A Survey of Environment Modeling, Synthesis, Evaluation, and Application

DGX agent

arXiv:2606.12191v1 Announce Type: cross Abstract: Environments serve as interactive systems for large language model (LLM) based agents across diverse scenarios and play a crucial role in driving the

agentsarxiv-cs-ai
11 Jun 2026
Agents

Agents All the Way Down; A Methodology for Building Custom AI Agents from Substrate to Production

DGX agent

arXiv:2606.11869v1 Announce Type: cross Abstract: Custom AI agents areagents that live inside their own application, talk to their own data and tools, enforce their own security boundaries, and carry

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

AI Coding Agents in Social Science: Methodologically Diverse, Empirically Consistent, Interpretively Vulnerable

DGX agent

arXiv:2606.11456v1 Announce Type: cross Abstract: The deployment of LLM-based agents in scientific analysis raises opposing concerns: that agents may reduce methodological diversity, or that they may

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

AI Researchers Must Help Lead Arms Control to Mitigate Military AI Risks

DGX agent

arXiv:2606.11533v1 Announce Type: cross Abstract: The advancement of AI capabilities compels researchers and the public to be more aware of its potential worldwide impact. A pressing near-term concern

safetyarxiv-cs-ai
11 Jun 2026
Hardware

AI4Land: Scalable Deep Learning for Global High-Resolution Land Use Reconstruction

DGX agent

arXiv:2606.11793v1 Announce Type: cross Abstract: Uncertainty in the terrestrial carbon cycle remains a major constraint in climate projections, partly driven by the uncertainties affecting the land s

hardwarearxiv-cs-ai
11 Jun 2026
Safety

ALIGNBEAM : Inference-Time Alignment Transfer via Cross-Vocabulary Logit Mixing

DGX agent

arXiv:2606.12342v1 Announce Type: cross Abstract: Domain fine-tuning degrades the safety of large language models: fine-tuned specialists readily comply with harmful prompts framed in domain language.

safetyarxiv-cs-ai
11 Jun 2026
Local Ai

Ambient Diffusion Policy: Imitation Learning from Suboptimal Data in Robotics

DGX agent

arXiv:2606.12365v1 Announce Type: cross Abstract: We propose Ambient Diffusion Policy, a simple and principled method for imitation learning from suboptimal data in robotics. High-quality, task-specif

local-aiarxiv-cs-ai
11 Jun 2026
Safety

An Ethical eValuation Agent (EeVA): Results of a Proof-of-Concept Test on a Prototype Agentic-like Workflow to Assist Ethical Deliberations

DGX agent

arXiv:2606.11218v1 Announce Type: cross Abstract: Ethical deliberation is often misunderstood as a search for single right or wrong answers, creating difficulties for non-ethically trained personnel w

safetyarxiv-cs-ai
11 Jun 2026
Research

An XAI View on Explainable ASP: Methods, Systems, and Perspectives

DGX agent

arXiv:2601.14764v2 Announce Type: replace Abstract: Answer Set Programming (ASP) is a popular declarative reasoning and problem solving approach in symbolic AI. Its rule-based formalism makes it inher

researcharxiv-cs-ai
11 Jun 2026
Model Releases

AnchorEdit: Maintaining Temporal Consistency in Multi-turn Image Editing via Causal Memory

DGX agent

arXiv:2606.11751v1 Announce Type: cross Abstract: Multi-turn image editing is essential for iterative design, yet current models often struggle with identity drift and error accumulation over successi

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

APEX: Automated Prompt Engineering eXpert with Dynamic Data Selection

DGX agent

arXiv:2606.11459v1 Announce Type: cross Abstract: Large Language Models are highly sensitive to prompt formulation, necessitating automatic prompt optimization to unlock their full potential. While ev

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

APPO: Agentic Procedural Policy Optimization

DGX agent

arXiv:2606.12384v1 Announce Type: cross Abstract: Recent advances in agentic Reinforcement Learning (RL) have substantially improved the multi-turn tool-use capabilities of large language model agents

safetyarxiv-cs-ai
11 Jun 2026
Safety

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning

DGX agent

arXiv:2606.11634v1 Announce Type: new Abstract: The rapid progress of reasoning and agentic large language models (LLMs) has increased the demand for long-context inference, but self-attention (SA) sc

safetyarxiv-cs-ai
11 Jun 2026
Research

Are LLMs Bad at Moral Reasoning?

DGX agent

arXiv:2606.11635v1 Announce Type: cross Abstract: For highly capable AI systems to operate safely in dynamic, open-ended environments, they must be able to identify, understand, and respond to moral r

researcharxiv-cs-ai
11 Jun 2026
Model Releases

ARGUS: Stacked Multi-View Identity Mosaic Injection for Subject-Preserving Video Generation

DGX agent

arXiv:2606.11670v1 Announce Type: cross Abstract: Subject-preserving video generation is not solved by frontal-face similarity alone: a generated person must remain recognizable across motion, large v

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

Artificial Intelligence in Ship Finance: Applications, Opportunities, and a Case Study in AI-Augmented Loan Origination

DGX agent

arXiv:2606.11238v1 Announce Type: cross Abstract: Ship finance is a data-intensive and document-heavy segment of asset-based lending, requiring the integration of financial, technical, contractual, an

agentsarxiv-cs-ai
11 Jun 2026
Agents

ATLAS: Active Theory Learning for Automated Science

DGX agent

arXiv:2606.12386v1 Announce Type: cross Abstract: Advancing scientific understanding through mechanistic modeling requires posing the right experimental questions to yield maximally informative data.

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

Atlas H&E-TME: Scalable AI-Based Tissue Profiling at Expert Pathologist-Level Accuracy

DGX agent

arXiv:2606.12346v1 Announce Type: cross Abstract: Hematoxylin and eosin (H&E) staining is the cornerstone of histopathology, yet scalable, quantitative analysis of H&E whole-slide images (WSIs) remain

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

Augmenting Molecular Language Models with Local n-gram Memory

DGX agent

arXiv:2606.12113v1 Announce Type: cross Abstract: Transformer-based language models for SMILES strings suffer from a locality gap: standard character-level tokenization fragments chemically meaningful

safetyarxiv-cs-ai
11 Jun 2026
Agents

Automated Creativity Evaluation of Language Models Across Open-Ended Tasks

DGX agent

arXiv:2606.11762v1 Announce Type: cross Abstract: Large language models (LLMs) have achieved remarkable progress in language understanding, reasoning, and generation, sparking growing interest in thei

agentsarxiv-cs-ai
11 Jun 2026
Agents

Automated Mediator for Human Negotiation: Pre-Mediation via a Structured LLM Pipeline

DGX agent

arXiv:2606.11379v1 Announce Type: new Abstract: Pre-mediation, the preparatory phase preceding direct human negotiation, plays a critical role in achieving mutually beneficial agreements, yet is often

agentsarxiv-cs-ai
11 Jun 2026
Safety

Automating Geometry-Intensive Compliance Checking in BIM: Graph-Based Semantic Reasoning Framework

DGX agent

arXiv:2606.12065v1 Announce Type: new Abstract: Automating compliance check for geometry-intensive regulations remains a significant technical bottleneck in Building Information Modeling (BIM), primar

safetyarxiv-cs-ai
11 Jun 2026
Safety

AutoMine Solution for AV2 2026 Scenario Mining Challenge

DGX agent

arXiv:2606.11874v1 Announce Type: new Abstract: With the development of autonomous driving systems, mining high-value, safety-critical, and planning-relevant scenarios from large-scale driving logs ha

safetyarxiv-cs-ai
11 Jun 2026
Research

Autoregressive Direct Preference Optimization

DGX agent

arXiv:2602.09533v2 Announce Type: replace Abstract: Direct preference optimization (DPO) has emerged as a promising approach for aligning large language models (LLMs) with human preferences. However,

researcharxiv-cs-ai
11 Jun 2026
Safety

AVIS: Adaptive Test-Time Scaling for Vision-Language Models

DGX agent

arXiv:2606.11576v1 Announce Type: cross Abstract: Modern Vision-Language Models (VLMs) benefit from chain-of-thought prompting and test-time scaling, but these gains often come with prohibitive infere

safetyarxiv-cs-ai
11 Jun 2026
Safety

Beyond representational alignment with brain-guided language models for robust reasoning

DGX agent

arXiv:2606.11893v1 Announce Type: cross Abstract: The correspondence between large language models (LLMs) and the neural mechanisms underlying human higher-order cognition remains insufficiently chara

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

BioDivergence: A Benchmark and Evaluation Framework for Hidden Contextual Contradictions in Biomedical Abstracts

DGX agent

arXiv:2606.11208v1 Announce Type: cross Abstract: Biomedical findings often seem to conflict across studies, but many of these differences are context-dependent rather than true contradictions. Variat

model-releasesarxiv-cs-ai
11 Jun 2026
Safety

Blind Dexterous Grasping via Real2Sim2Real Tactile Policy Learning

DGX agent

arXiv:2606.11767v1 Announce Type: cross Abstract: Blind grasping with a dexterous hand is a crucial manipulation capability. Nevertheless, learning such tactile-only policies for real robots remains c

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Bridging the Morphology Gap: Adapting VLA Models to Dexterous Manipulation via Intent-Conditioned Fine-Tuning

DGX agent

arXiv:2606.12109v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable zero-shot generalization in robotic manipulation, yet the vast majority of pre-traine

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Calibration Drift Under Reasoning: How Chain-of-Thought Budgets Induce Overconfidence in Large Language Models

DGX agent

arXiv:2606.11211v1 Announce Type: cross Abstract: The ability of large language models (LLMs) to express calibrated uncertainty is important for safe deployment. Chain-of-thought (CoT) reasoning is wi

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Can AI Agents Synthesize Scientific Conclusions?

DGX agent

arXiv:2606.11337v1 Announce Type: new Abstract: Scientific AI agents increasingly retrieve evidence, reason across sources, and synthesize conclusions used in consequential decisions. Yet, their abili

model-releasesarxiv-cs-ai
11 Jun 2026
Local Ai

Can Open-Source LLM Agents Replace Static Application Security Testing Tools? An Empirical Assessment

DGX agent

arXiv:2606.11672v1 Announce Type: cross Abstract: This paper explores the value of agentic AI tools for cybersecurity purposes. We evaluate the efficacy of a general-purpose GenAI Large Language Model

local-aiarxiv-cs-ai
11 Jun 2026
← Previous
1…162163164165166…448
Next →