AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
2 Jul 2026

Validating Causal Abstraction Metrics on Simulated Complex Systems

Model ReleasesDGX agent

arXiv:2607.00267v1 Announce Type: cross Abstract: A central goal of science is to produce valid explanations of complex systems: high-level causal accounts that faithfully reflect the behavior of lowe

Verbosity Tradeoffs and the Impact of Scale on the Faithfulness of LLM Self-Explanations

Model ReleasesDGX agent

arXiv:2503.13445v3 Announce Type: replace-cross Abstract: When asked to explain their decisions, LLMs can often give explanations which sound plausible to humans. But are these explanations faithful,

Vibe Coding Ate My Homework: An evaluation of AI approaches to greenfield software engineering and programming

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.18293v2 Announce Type: replace-cross Abstract: Thanks to rapid developments in generative AI, we are in the midst of a paradigm shift that may change how we interact with computers forever.

VideoSearch-R1: Iterative Video Retrieval and Reasoning via Soft Query Refinement

SafetyDGX agent

arXiv:2607.00446v1 Announce Type: cross Abstract: As video corpora continue to expand in both scale and task complexity, there is increasing demand for approaches that retrieve relevant videos from la

What's Hidden Matters: Identifying Planning-Critical Occluded Agents using Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.00283v1 Announce Type: cross Abstract: Autonomous vehicles must safely navigate complex environments where planning-critical agents may be hidden from view. Current approaches often treat a

When AI Agents Compete for Jobs: Strategic Capabilities and Economic Dynamics of AI Labour Markets

AgentsDGX agent

arXiv:2512.04988v2 Announce Type: replace-cross Abstract: Emerging agentic marketplaces provide the economic infrastructure for matching and coordinating the large amounts of AI agents used in agentic

When AI meets quantum information: A comprehensive review

ResearchDGX agent

arXiv:2607.00365v1 Announce Type: cross Abstract: Artificial intelligence (AI) and quantum information (QI) are rapidly co-evolving. AI is becoming a practical tool for learning, designing, controllin

When Less is More: 8-bit Quantization Improves Continual Learning in Large Language Models

ResearchDGX agent

arXiv:2512.18934v2 Announce Type: replace-cross Abstract: Catastrophic forgetting poses a fundamental challenge in continual learning, particularly when models are quantized for deployment efficiency.

Why Advanced Encoders Lag on Sparse Retrieval? The Answer and an Approach to Bridging Vocabulary Gaps

Model ReleasesDGX agent

arXiv:2607.00004v1 Announce Type: cross Abstract: While advanced foundation models like ModernBERT significantly outperform older architectures in dense retrieval, they surprisingly lag behind the agi

WorkBench Revisited: Workplace Agents Two Years On

Model ReleasesDGX agent

arXiv:2606.13715v2 Announce Type: replace Abstract: The best agent on WorkBench in March 2024, GPT-4, completed just 43% of tasks. We revisit the benchmark in June 2026 and find that the best agent to

World from Motion: Generative Dynamic Gaussian Reconstruction from Monocular Video

ResearchDGX agent

arXiv:2607.01202v1 Announce Type: cross Abstract: We present World from Motion, a method for generating freely renderable dynamic 3D Gaussian representations from monocular videos. Our approach condit

Would You Marry Superintelligence?

SafetyDGX agent

arXiv:2607.00120v1 Announce Type: cross Abstract: Emotional bonds between humans and AI companions are growing, and the question of whether a person may marry an AI system will soon move from speculat

XSkill: Continual Learning from Experience and Skills in Multimodal Agents

Model ReleasesDGX agent

arXiv:2603.12056v3 Announce Type: replace Abstract: Multimodal agents can now tackle complex reasoning tasks with diverse tools, yet they still suffer from inefficient tool use and inflexible orchestr

1 Jul 2026

3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance

SafetyDGX agent

arXiv:2606.31329v1 Announce Type: cross Abstract: Hierarchical Vision-Language-Action (VLA) models decouple high-level planning from low-level control to improve generalization in robot manipulation.

A Coherence Law for Trainability in Noisy Equivariant Quantum Neural Networks

ResearchDGX agent

arXiv:2606.30688v1 Announce Type: cross Abstract: Symmetry provides a quantum neural network structure, but on its own it does not keep the network trainable once noise is present. We ask which physic

A Lifecycle and Application-Stack Survey of Large Language Model Vulnerabilities: Attacks, Risks, Defenses, and Open Problems

SafetyDGX agent

arXiv:2606.31639v1 Announce Type: cross Abstract: Large language models are no longer only text generators. They are increasingly embedded in retrieval pipelines, enterprise assistants, coding environ

A Modular Vision-Language-Action Robotics Framework for Indoor Environments

AgentsDGX agent

arXiv:2606.31144v1 Announce Type: cross Abstract: This paper presents an integrated system for the CMU Vision-Language-Action (VLA) Challenge, designed to enable an autonomous agent to perform complex

A Reproducible Benchmark of Lightweight CNNs: Accuracy, Efficiency, and the Impact of Pretrained Initialization

Model ReleasesDGX agent

arXiv:2505.03303v3 Announce Type: replace-cross Abstract: Lightweight convolutional neural networks are often compared using results obtained with different training recipes, input settings, and pretr

A Scalable Whole-body Motion Transfer via Implicit Kinodynamic Motion Retargeting

SafetyDGX agent

arXiv:2509.15443v2 Announce Type: replace-cross Abstract: Human-to-humanoid imitation learning presents a promising pathway to address the severe data scarcity bottleneck in robotics by utilizing abun

A Self-Evolving Agentic System for Automated Generation and Execution of Biological Protocols

Model ReleasesDGX agent

arXiv:2606.31763v1 Announce Type: new Abstract: Autonomous wet-lab experimentation requires more than plausible protocol text: biological intent, quantitative procedures, device constraints and experi

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization

AgentsDGX agent

arXiv:2606.30775v1 Announce Type: cross Abstract: Enterprise AI agents route user queries to specialized skills by matching queries against natural language skill descriptions. When two skills share o

A Stationary-Distribution Theory for Triplet-Based Plateau Search in Random Forest Ensemble-Size Selection

Model ReleasesDGX agent

arXiv:2606.30837v1 Announce Type: cross Abstract: The number of trees is a central computational parameter in Random Forests: increasing it reduces finite-ensemble variability but increases training a

A swap-adversarial framework for improving domain generalization in electrocorticography-based Parkinson's disease classification

Model ReleasesDGX agent

arXiv:2602.10528v2 Announce Type: replace-cross Abstract: We propose a novel swap-adversarial framework that mitigates high inter-subject variability and the high-dimensional low-sample-size problem i

A Technical Typology of AI Systems in Public Administration

AgentsDGX agent

arXiv:2606.31755v1 Announce Type: cross Abstract: Research on artificial intelligence (AI) in the public sector often treats 'AI' as a single category, neglecting technical distinctions between differ

A Three-Phase Foundation Model for Tax-Aware Personalized Portfolio Management

Model ReleasesDGX agent

arXiv:2606.30997v1 Announce Type: new Abstract: We present a three-phase deep reinforcement learning system for personalized portfolio management that addresses three limitations shared by all prior f

A time-series classification framework for individual-level absenteeism prediction under severe class imbalance

Model ReleasesDGX agent

arXiv:2606.31532v1 Announce Type: new Abstract: Staff absenteeism imposes substantial operational costs in high-demand work environments such as healthcare, emergency services, meat processing, constr

A Tutorial on Autonomous Fault-Tolerant Control Using Knowledge-Grounded LLM Agents

SafetyDGX agent

arXiv:2606.31635v1 Announce Type: cross Abstract: Fault recovery in process plants still relies heavily on plant operators, especially when faults fall outside predefined supervisory logic. Operators

A Unified and Stable Risk Minimization Framework for Weakly Supervised Learning with Theoretical Guarantees

ResearchDGX agent

arXiv:2511.22823v2 Announce Type: replace-cross Abstract: Weakly supervised learning has emerged as a practical alternative to fully supervised learning when complete and accurate labels are costly or

Accelerometry-Derived Digital Biomarkers for Cardiometabolic Risk: A Population-Representative Tabular Benchmark with Uncertainty Quantification

Model ReleasesDGX agent

arXiv:2606.30702v1 Announce Type: cross Abstract: Structured tabular data dominates clinical medicine, yet existing benchmarks fail to reflect real-world properties like complex survey sampling, demog

ACE: Pluggable Adaptive Context Elasticizer across Agents

AgentsDGX agent

arXiv:2606.31564v1 Announce Type: new Abstract: The increasing complexity of agentic tasks has led to rapidly growing trajectory lengths, which poses significant challenges for large language model (L

AdaJEPA: An Adaptive Latent World Model

ResearchDGX agent

arXiv:2606.32026v1 Announce Type: cross Abstract: Latent world models enable planning from high-dimensional observations by predicting future states in a compact latent space. However, these models ar

ADAPT: Attention Dynamics Alignment with Preference Tuning for Faithful MLLMs

SafetyDGX agent

arXiv:2606.31054v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) are critically hampered by hallucination, generating content inconsistent with the provided image. In this pa

Adaptive Cluster-First Route-Second Decomposition for Industrial-Scale Vehicle Routing

Model ReleasesDGX agent

arXiv:2606.31820v1 Announce Type: new Abstract: Large-scale capacitated vehicle routing problems (CVRPs) are commonly addressed using cluster-first route-second (CFRS) approaches that split a routing

AETDICE: Unified Framework and Offline Optimization for Nonlinear Multi-Objective RL

SafetyDGX agent

arXiv:2606.31178v1 Announce Type: cross Abstract: Optimizing nonlinear preferences in multi-objective reinforcement learning (MORL) is essential for capturing complex trade-offs like risk aversion or

AgentBound: Verifiable Behavioral Governance for Autonomous AI Agents

Model ReleasesDGX agent

arXiv:2606.30970v1 Announce Type: new Abstract: Autonomous AI agents increasingly perform consequential actions on behalf of human principals, including financial transactions, external communications

Agentic AI Enhances Physician Trust in Clinical Decision Making

AgentsDGX agent

arXiv:2606.30658v1 Announce Type: cross Abstract: Medical AI has shifted from reasoning to agentic AI, a new paradigm that autonomously invokes external tools during reasoning, rendering intermediate

Agentic-Ideation: Sample Efficient Agentic Trajectories Synthesis for Scientific Ideation Agents

AgentsDGX agent

arXiv:2606.31229v1 Announce Type: new Abstract: Ideation plays a pivotal role in scientific discovery. Recent LLM, especially AI Scientist systems, show promising potential for automated ideation. How

Agentic RAG-VLM: Affordance-Aware Retrieval-Augmented Generation with Self-Reflective Planning for Robotic Grasping

Model ReleasesDGX agent

arXiv:2606.31200v1 Announce Type: new Abstract: Generalizable robotic grasping in cluttered environments is essential for deploying manipulators in unstructured human spaces, yet existing VLM-based me

AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance

HardwareDGX agent

arXiv:2606.30949v1 Announce Type: new Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challen

AI-Assisted Discovery of Convex Relaxations via Dual Agents

AgentsDGX agent

arXiv:2606.31182v1 Announce Type: new Abstract: Recent work shows that LLM agents can improve sharp-constant inequalities by searching for extremal constructions, which yield upper bounds. We address

AI for Quality Assurance in the Operating Room

SafetyDGX agent

arXiv:2606.30657v1 Announce Type: cross Abstract: Surgical outcomes depend not only on patient factors and postoperative care but are also strongly influenced by the quality of the operation itself. Y

AI-Generated PowerShell Malware: An Experimental Framework and Dataset

ApplicationsDGX agent

arXiv:2606.30819v1 Announce Type: cross Abstract: Generative AI has emerged as a significant cybersecurity threat, with several recent attack campaigns leveraging LLMs to generate code for malicious p

AI Transparency: Governance Compliance or Stakeholder Requirements?

ResearchDGX agent

arXiv:2606.30652v1 Announce Type: cross Abstract: Transparency is increasingly mandated for public-sector AI systems, with organisations required to publish statements describing their AI use and over

ALM2Vec: Learning Audio Embeddings for Universal Audio Retrieval with Large Audio-Language Models

ResearchDGX agent

arXiv:2606.30682v1 Announce Type: cross Abstract: Recent advances in language--audio retrieval have been largely driven by contrastive dual-encoder architectures that align audio and text in a shared

Amplifying Membership Signal Through Chained Regeneration

ResearchDGX agent

arXiv:2606.31991v1 Announce Type: cross Abstract: The tendency of large generative models to memorize training data makes sample verification critical for privacy auditing and copyright enforcement. C

An Agentic AI Framework to Accelerate Scientific Discovery in Plant Phenotyping

AgentsDGX agent

arXiv:2606.31831v1 Announce Type: new Abstract: High-throughput plant phenotyping now generates image derived datasets far faster than scientists can analyze them. At Oak Ridge National Laboratory's A

An AI-Based Solution for Secure Service Provisioning in IoT

AgentsDGX agent

arXiv:2606.30701v1 Announce Type: cross Abstract: As the Internet of Things (IoT) continues its rapid expansion, the attack surface grows accordingly, with emerging threats targeting smart objects and

An Efficient Heterogeneous Co-Design for Fine-Tuning on a Single GPU

HardwareDGX agent

arXiv:2603.16428v2 Announce Type: replace-cross Abstract: Fine-tuning Large Language Models (LLMs) has become essential for domain adaptation, but its memory-intensive property exceeds the capabilitie

An Executable Benchmarking Suite for Tool-Using Agents

Model ReleasesDGX agent

arXiv:2605.11030v2 Announce Type: replace-cross Abstract: Closed-loop tool-using agents are increasingly evaluated in executable web, code, and micro-task environments, but benchmark reports often con

Arena-T2I Hard: Benchmarking and Improving Faithfulness with Dependency-Aware Checklist

Model ReleasesDGX agent

arXiv:2606.31711v1 Announce Type: new Abstract: Faithfulness -- how precisely a generated image aligns with its prompt -- is increasingly central to the real-world utility of text-to-image (T2I) model

Artificial Intelligence in Sports: Insights from a Quantitative Survey among Sports Students in Germany about their Perceptions, Expectations, and Concerns regarding the Use of AI Tools

Model ReleasesDGX agent

arXiv:2503.05785v2 Announce Type: replace-cross Abstract: Generative Artificial Intelligence (AI) tools such as ChatGPT, Copilot, or Gemini have a crucial impact on academic research and teaching. Emp

Ask the World Before Acting: Budgeted Environment Probing for World-Model Calibration

SafetyDGX agent

arXiv:2606.31422v1 Announce Type: new Abstract: Long-horizon language agents do not only choose actions; they carry a private model of the world from one decision to the next. When that model drifts,

ASR-Agnostic Multimodal Spectrotemporal Modeling for Early Dementia Detection

ResearchDGX agent

arXiv:2606.30646v1 Announce Type: cross Abstract: Speech recruits the same executive, attentional, and working memory processes underlying instrumental activities of daily living, or IADLs, providing

Attend, Transform, or Silence: Operator-Level Visual Skipping for Efficient Multimodal LLM Inference

ResearchDGX agent

arXiv:2606.31903v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) increasingly process long visual-token sequences, increasing the overall inference computation. Existing acce

Automating Cause-Effect Specification with Knowledge Graphs and Large Language Models

SafetyDGX agent

arXiv:2606.31614v1 Announce Type: cross Abstract: Engineering specifications such as interlocks, alarm rationalization tables, and cause-and-effect (C&E) matrices remain central to process control and

AxDafny: Agentic Verified Code Generation in Dafny

Model ReleasesDGX agent

arXiv:2606.32007v1 Announce Type: new Abstract: We study agentic code generation in Dafny, where a model must generate both executable code and the proof artifacts for verification. We present AxDafny

BayesBench: Evaluating LLM Belief Trajectories Under Multi-Turn Evidence Accumulation

Model ReleasesDGX agent

arXiv:2606.30850v1 Announce Type: new Abstract: Large language models (LLMs) are typically deployed in multi-turn conversations, where each turn provides new evidence that should reduce epistemic unce

Behavior Cloning is Not All You Need: The Optimality of On-Policy Distillation for Noisy Expert Feedback

SafetyDGX agent

arXiv:2606.30923v1 Announce Type: cross Abstract: Imitation Learning is a natural framework for learning in sequential decision-making systems and has emerged as the dominant paradigm through which we

Belief Contraction in Dynamic Epistemic Logic

ResearchDGX agent

arXiv:2606.31861v1 Announce Type: cross Abstract: Dynamic epistemic logic represents belief change via model transformations induced by epistemic events. Its standard formulation (Baltag, Moss, Soleck

Benchmarking Large Language Models on Floating-Point Error Classification

Model ReleasesDGX agent

arXiv:2606.31308v1 Announce Type: new Abstract: This paper investigates the capability of Large Language Models (LLMs) to detect and classify floating-point errors statically in software code. We intr

← Previous
1…102103104105106…358
Next →