AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
7 Jul 2026

VISTA: Auditing Semantic Divergence in Vision-Language Models

ResearchDGX agent

arXiv:2607.02995v1 Announce Type: cross Abstract: Vision-language models can exhibit visual concept-conditioned divergence: given images containing demographic features, corporate logos, or ideologica

VLA Grounder: Language-Conditioning Space Optimization for Black-Box VLA Models

SafetyDGX agent

arXiv:2607.04517v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are commonly treated as end-to-end action policies conditioned on natural-language task descriptions. In practice, h

Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

Model ReleasesDGX agent

arXiv:2511.15684v2 Announce Type: replace-cross Abstract: Foundation models have transformed machine learning for language and vision, but achieving comparable impact in physical simulation remains a


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Wan-Streamer v0.2: Higher Resolution, Same Latency

HardwareDGX agent

arXiv:2607.04443v1 Announce Type: cross Abstract: We present Wan-Streamer v0.2, a latency-preserving upgrade of the native-streaming, end-to-end audio-visual interaction model. v0.2 keeps the v0.1 mod

Wasserstein Residuals: Learning Gradient Flows from Population Dynamics

ResearchDGX agent

arXiv:2607.04738v1 Announce Type: cross Abstract: Reconstructing population dynamics is a central problem in the physical and data sciences. Often, the dynamics are modeled as a Wasserstein gradient f

Wavelet Scattering Transform for Interpretable Schizophrenia Biomarker Discovery and Classification from Resting-State EEG

ResearchDGX agent

arXiv:2607.05282v1 Announce Type: cross Abstract: Schizophrenia is a debilitating neuropsychiatric disorder characterized by profound cortical network dysregulation, for which objective, clinically tr

Weak-to-Strong Generalization via Direct On-Policy Distillation

SafetyDGX agent

arXiv:2607.05394v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a powerful recipe for improving language-model reasoning, but it is expensive to repeat on ev

Web-CogReasoner: Towards Multimodal Knowledge-Induced Cognitive Reasoning for Web Agents

AgentsDGX agent

arXiv:2508.01858v3 Announce Type: replace-cross Abstract: Multimodal large-scale models have significantly advanced the development of web agents, enabling perception and interaction with digital envi

What Does a Discrete Diffusion Model Learn?

TutorialsDGX agent

arXiv:2607.05381v1 Announce Type: cross Abstract: What does a discrete diffusion model learn: a denoiser, a score ratio, or a bridge plug-in predictor? At the level of jump rates, these are one object

What is Left for Us? Second Scholarship Against the Degradation of Research by AI

ApplicationsDGX agent

arXiv:2607.04049v1 Announce Type: new Abstract: We argue that generative AI can degrade research by eroding the very practices through which scholarly judgement is formed and academic trust is built.

When Aggregate Alignment Misleads: Auditing Policy Repair Without Per-State Expert Actions

Model ReleasesDGX agent

arXiv:2607.03386v1 Announce Type: new Abstract: Agentic AI systems are increasingly used to edit, refine, and repair decision policies, but evaluating these edits is difficult when per-state expert ac

When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Agents

Model ReleasesDGX agent

arXiv:2607.05189v1 Announce Type: cross Abstract: Persistent personal agents combine long-term memory with access to users' external environments, enabling personalized foreground assistance and proac

When Does Small Data Work? Accuracy and Efficiency Trade-offs Between Tabular Foundation Models and Conventional Methods for Crowd-State Classification at Hajj and Umrah

ResearchDGX agent

arXiv:2607.04013v1 Announce Type: cross Abstract: Learning from few labeled examples is a central challenge in tabular machine learning, and it becomes the binding constraint in domains where labeling

When is a System Discoverable from Data? Discovery Requires Chaos

Model ReleasesDGX agent

arXiv:2511.08860v2 Announce Type: replace-cross Abstract: The deep learning revolution has spurred a rise in advances of using AI in sciences. Within physical sciences the main focus has been on disco

When Rubrics Fail: Error Enumeration as Reward in Reference-Free RL Post-Training for Virtual Try-On

Model ReleasesDGX agent

arXiv:2603.05659v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) and Rubrics as Rewards (RaR) have driven strong gains in domains with clear correctness

When Simpler Is Better: Evaluating Translation Pipelines for Medieval Latin Manuscripts

Model ReleasesDGX agent

arXiv:2607.03836v1 Announce Type: cross Abstract: Despite remarkable progress in machine translation, Vision Language Models (VLMs) struggle on historical manuscripts, a domain that stresses core Natu

Where do LLMs Fall Short in CBT-Guided Affective Reasoning?

ResearchDGX agent

arXiv:2607.02885v1 Announce Type: cross Abstract: Cognitive Behavioral Therapy (CBT) provides a structured framework for understanding a user's mental state by examining the interaction between cognit

Which Algorithm Specification Formats Help Language Models Implement Machine Learning Algorithms?

Model ReleasesDGX agent

arXiv:2607.03158v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to implement algorithms from research manuscripts, but papers often leave implementation choices im

Why Pure Reasoning is Not Enough: Nature as the Source of Mathematical Innovation

ResearchDGX agent

arXiv:2607.04505v1 Announce Type: new Abstract: We advance the hypothesis that human mathematical reasoning, constrained by both the undecidability and the computational intractability of even modest

Why3-py: A Tool for Formal Verification of Hypothesis Testing and Meta-Analysis in Python

ResearchDGX agent

arXiv:2607.03951v1 Announce Type: cross Abstract: The reproducibility crisis in scientific research has received widespread recognition, thereby increasing the importance of meta-analyses that integra

Worldscape-MoE: A Unified Mixture-of-Experts World Model for Scalable Heterogeneous Action Control

ResearchDGX agent

arXiv:2607.03964v1 Announce Type: cross Abstract: World models are rapidly becoming a core infrastructure for embodied intelligence and interactive agents: they provide controllable simulators in whic

Your Agent's Memories Are Not Its Own: Forged Reasoning Attacks on LLM Agent Memory and Defenses

AgentsDGX agent

arXiv:2607.05029v1 Announce Type: cross Abstract: Persistent memory has enabled large language model (LLM) agents to store factual knowledge, prior decisions, reasoning histories, tool usage informati

3 Jul 2026

A Dual-Helix Governance Approach Towards Reliable Agentic Artificial Intelligence for WebGIS Development

AgentsDGX agent

arXiv:2603.04390v2 Announce Type: replace Abstract: WebGIS development requires consistency, yet agentic AI often fails due to LLM context constraints, forgetting, stochasticity, instruction failure,

A General Neural Backbone for Mixed-Integer Linear Optimization via Dual Attention

Local AiDGX agent

arXiv:2601.04509v2 Announce Type: replace Abstract: Mixed-integer linear programming (MILP) is a foundational framework for combinatorial optimization across science and engineering, but remains hard

A Hippocampus for Linear Attention: An Exact Memory for What the Recurrent State Forgets

ResearchDGX agent

arXiv:2607.02303v1 Announce Type: new Abstract: Linear-attention and state-space language models compress the prefix into a fixed-size recurrent state, yielding O(1) memory at the cost of a lossy exac

A Multi-Branch Hierarchy-Aware Framework for Heterogeneous Audio Classification

ResearchDGX agent

arXiv:2607.01974v1 Announce Type: cross Abstract: This technical report describes our system for Task 1 of the DCASE 2026 Challenge, which aims to classify heterogeneous audio recordings according to

A Practice Auditing Framework for Large Language Model Use: Collective Empiricism, Pseudo-Rational Cognition, and Governance of AI-Generated Content

AgentsDGX agent

arXiv:2607.01248v1 Announce Type: cross Abstract: Large language models are increasingly used for knowledge acquisition, code generation, academic writing, and agent-based automation. In these setting

A rubric-based controlled comparison of frontier language models on expert-authored clinical reasoning tasks

Model ReleasesDGX agent

arXiv:2607.02175v1 Announce Type: new Abstract: Multiple-choice medical benchmarks are increasingly saturated, and recent rubric-based evaluations such as HealthBench have shown that open-ended clinic

A-TMA: Decoupling State-Aware Memory Failures in Long-Term Agent Memory

Model ReleasesDGX agent

arXiv:2607.01935v1 Announce Type: new Abstract: Long term memory lets LLM agents act as persistent assistants, but user facts change. A useful memory system must know what is true now, what used to be

A^{2}utoLPBench: An Auto-Generated, Agent-Friendly LP Benchmark via Inverse-KKT Construction

Model ReleasesDGX agent

arXiv:2607.02141v1 Announce Type: new Abstract: Most LP-from-text benchmarks are static datasets of word problems written and labeled by hand. Once such a dataset is released, its size is fixed, its d

ACID: Action Consistency via Inverse Dynamics for Planning with World Models

ResearchDGX agent

arXiv:2607.02403v1 Announce Type: cross Abstract: Decision-time planning with action-conditioned world models has become a popular paradigm for embodied control. However, the standard planning cost ju

Activation Steering for Aligned Open-ended Generation without Sacrificing Coherence

Model ReleasesDGX agent

arXiv:2604.08169v2 Announce Type: replace Abstract: Alignment in LLMs is more brittle than commonly assumed: misalignment can be induced by adversarial prompts, benign fine-tuning, emergent misalignme

Actual causality in fault trees

ResearchDGX agent

arXiv:2607.01840v1 Announce Type: new Abstract: Fault trees are a widely used as effective risk models for complex systems, answering the question 'what can go wrong?', especially through minimal cut

Adaptive Batch Sizes Using Non-Euclidean Gradient Noise Scales for Stochastic Sign and Spectral Descent

Model ReleasesDGX agent

arXiv:2602.03001v2 Announce Type: replace-cross Abstract: To maximize hardware utilization, modern machine learning systems typically employ large constant or manually tuned batch size schedules, rely

Adaptive Companionship for Group-Following Robots: Handling Dynamically Changing Group Formations

SafetyDGX agent

arXiv:2607.01287v1 Announce Type: cross Abstract: Accompanying a group of humans is an essential aspect of developing human-like social cognition in robots. However, human groups typically do not foll

Adaptive Contracts for Cost-Effective AI Delegation

ResearchDGX agent

arXiv:2603.17212v2 Announce Type: replace-cross Abstract: When organizations delegate text generation tasks to AI providers via pay-for-performance contracts, expected payments rise when evaluation is

ADMC: Attention-based Diffusion Model for Missing Modalities Feature Completion

ResearchDGX agent

arXiv:2507.05624v2 Announce Type: replace Abstract: Multimodal emotion and intent recognition is essential for automated human-computer interaction, It aims to analyze users' speech, text, and visual

Adoption and Impact of Command-Line AI Coding Agents: A Study of Microsoft's Early 2026 Rollout of Claude Code and GitHub Copilot CLI

Model ReleasesDGX agent

arXiv:2607.01418v1 Announce Type: cross Abstract: Organizations rolling out agentic command line tools like Anthropic's Claude Code and GitHub's Copilot CLI need to know who will try them, who will ke

ADVENT: LLM-Driven Automatic Predicate Invention for ILP

TutorialsDGX agent

arXiv:2607.01585v1 Announce Type: cross Abstract: Predicate invention (PI), the creation of new predicates to extend the hypothesis space, remains a critical bottleneck in Inductive Logic Programming

Agent4cs: A Multi-agent System for Code Summarization in Large Hierarchical Codebases

Model ReleasesDGX agent

arXiv:2607.01425v1 Announce Type: new Abstract: Understanding large, complex codebases, especially those with obfuscated structures and incomplete documentation, remains a significant challenge. Exist

AgenticDataBench: A Comprehensive Benchmark for Data Agents

Model ReleasesDGX agent

arXiv:2607.01647v1 Announce Type: cross Abstract: Data science aims to derive actionable insights from heterogeneous raw data, unlocking the value of the massive amounts of data generated in modern so

AgenticSTS: A Bounded-Memory Testbed for Long-Horizon LLM Agents

Model ReleasesDGX agent

arXiv:2607.02255v1 Announce Type: new Abstract: Memory for a long-horizon LLM agent is a contract about what each future decision is allowed to see. The simplest contract appends past observations, to

AI Assistance for Human Review of Default Judgments

ApplicationsDGX agent

arXiv:2607.01256v1 Announce Type: cross Abstract: Overwhelmed courts in the United States review millions of default judgments each year. Unfortunately, such manual reviews are time-consuming and pron

AI-enabled gravitational-waves searches for binary neutron stars at optimal sensitivity

HardwareDGX agent

arXiv:2607.01372v1 Announce Type: cross Abstract: Gravitational Waves (GWs) represent the newest window of astronomy, furthering our understanding of compact objects like black holes and neutron stars

AI Virtue: What is 'Good' Knowledge in the Age of Artificial Intelligence?

ResearchDGX agent

arXiv:2607.01776v1 Announce Type: cross Abstract: In the age of AI, what will be good knowledge? This article, which is accepted and forthcoming in a special issue of Modern Fiction Studies on 'Cultur

AIriskEval-edu: New Dataset for Risk Assessment in AI-mediated K-12 Educational Explanations

Model ReleasesDGX agent

arXiv:2607.01934v1 Announce Type: cross Abstract: This work introduces AIriskEval-edu-db2, a new dataset designed to train and evaluate auditors based on LLMs for an explainable pedagogical risk asses

Algebraic Model Counting for Global Analysis of Optimal Decision Trees

SafetyDGX agent

arXiv:2607.02069v1 Announce Type: new Abstract: Ensuring model reliability in Explainable AI requires a global assessment of the hypothesis space. We propose a formal framework for the exhaustive anal

An Efficient vLLM-Based Inference Pipeline for Unified Audio Understanding and Generation

HardwareDGX agent

arXiv:2607.02119v1 Announce Type: cross Abstract: While Large Multimodal Models excel in comprehension, high-throughput inference engines lack native support for multimodal generation. This is severe

An Exploratory Study on LLM-Generated Code and Comments in Code Repositories

ResearchDGX agent

arXiv:2607.01867v1 Announce Type: cross Abstract: The use of LLMs in software development has become increasingly widespread on tasks such as code generation and summarization. Reports from large tech

An Isotropic Approach to Efficient Uncertainty Quantification with Gradient Norms

Model ReleasesDGX agent

arXiv:2603.29466v2 Announce Type: replace-cross Abstract: Existing methods for quantifying predictive uncertainty in neural networks are either computationally intractable for large language models or

AnyGroundBench: A Specialized-Domain Benchmark for Video Grounding in Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.02269v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated immense promise in Spatio-Temporal Video Grounding (STVG). However, current evaluation protocols are l

Aria: An Agent For Retrieval and Iterative Auto-Formalization via Dependency Graph

AgentsDGX agent

arXiv:2510.04520v2 Announce Type: replace Abstract: Accurate auto-formalization of theorem statements is essential for advancing automated discovery and verification of research-level mathematics, yet

ART for Diffusion Sampling: Continuous-Time Control and Actor-Critic Learning

SafetyDGX agent

arXiv:2607.02137v1 Announce Type: cross Abstract: We study timestep allocation for score-based diffusion sampling, where a learned reverse-time dynamics is discretized on a finite grid. Uniform and ha

Artificial Intelligence-Enabled Accounting Information Systems and Fraud Detection in Nigeria's Financial Services Sector: The Moderating Role of Natural Language Processing

ResearchDGX agent

arXiv:2607.01257v1 Announce Type: cross Abstract: The rapid digitalisation of financial systems has improved operational efficiency and financial inclusion while simultaneously increasing exposure to

Assessing VLM Reliability for Medical Image Quality Evaluation Under Corruption and Bias

Model ReleasesDGX agent

arXiv:2607.01973v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly applied in medical tasks such as pathology description, report generation, and visual question answerin

Atomic Task Graph: A Unified Framework for Agentic Planning and Execution

AgentsDGX agent

arXiv:2607.01942v1 Announce Type: new Abstract: LLM-based agents have shown strong potential for solving complex multi-step tasks, yet existing performance improvements often rely on either scaling to

Auto-FL-Research: Agentic Search for Federated Learning Algorithms

Local AiDGX agent

arXiv:2607.01366v1 Announce Type: new Abstract: Federated learning (FL) research often depends on many small but consequential algorithmic choices: optimizer variants, server aggregation rules, local

Automated grading of Linux/bash examinations using large language models: a four-level cognitive taxonomy approach

Model ReleasesDGX agent

arXiv:2607.02432v1 Announce Type: new Abstract: Scalable and reliable grading of command-line examinations remains a challenge in computing education, where rising enrolments make manual marking diffi

Autonomous discovery of traffic laws with AI traffic scientists

AgentsDGX agent

arXiv:2607.01639v1 Announce Type: new Abstract: Universal traffic laws describe recurrent patterns in congestion, mobility and driving behavior across cities, providing a scientific basis for transpor

Behind the Refusal: Determining Guardrail Activation via Behavioral Monitoring

SafetyDGX agent

arXiv:2607.02121v1 Announce Type: cross Abstract: As Large Language Models (LLMs) and agentic systems become integrated into real-world applications, ensuring their safety and security is critical. Gu

← Previous
1…9394959697…358
Next →