AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
9 Jul 2026

Understanding Two-Layer Neural Networks with Smooth Activation Functions

ResearchDGX agent

arXiv:2507.14177v2 Announce Type: replace-cross Abstract: This paper aims to understand the training solution, which is obtained by the back-propagation algorithm, of two-layer neural networks whose h

Validate the Dream Before You Trust Its Verdict: Admissibility for World-Model Simulators

SafetyDGX agent

arXiv:2607.07196v1 Announce Type: cross Abstract: Across robotics, World Models (WMs) are increasingly used to evaluate action policies by simulating the consequences of actions in an imagined world,

Vision Foundation Models in Radiology: A Scoping Review of Data, Methodology, Evaluation and Clinical Translation

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.07219v1 Announce Type: cross Abstract: Vision foundation models (VFMs) are increasingly being developed for radiological imaging, yet their definition, development and evaluation remain het

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review

ResearchDGX agent

arXiv:2607.06706v1 Announce Type: cross Abstract: Vision Language Action (VLA) models unify visual perception, natural-language understanding, and action generation within a single foundation model, a

VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

ResearchDGX agent

arXiv:2507.05116v5 Announce Type: replace-cross Abstract: Recent large-scale Vision Language Action (VLA) models have shown superior performance in robotic manipulation tasks guided by natural languag

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time

ResearchDGX agent

arXiv:2607.06988v1 Announce Type: cross Abstract: Steering robot foundation models (RFMs) toward new task variants or user-preferred behaviors remains challenging, often requiring additional robot dem

What Predicts Correctness in Text-to-SQL? A Selective-Prediction Study

Model ReleasesDGX agent

arXiv:2607.06799v1 Announce Type: cross Abstract: Evaluating uncertainty in AI-generated SQL queries requires estimating whether a query is correct, where correct means it executes to the same result

When Agents Go Rogue: Activation-Based Detection of Malicious Behaviors in Multi-Agent Systems

Local AiDGX agent

arXiv:2607.06807v1 Announce Type: cross Abstract: While enabling effective collaboration on complex tasks, LLM-based Multi-Agent Systems (MAS) face critical security challenges due to vulnerabilities

When Agents Remember Too Much: Memory Poisoning Attacks on Large Language Model Agents

SafetyDGX agent

arXiv:2607.06595v1 Announce Type: cross Abstract: Personal AI agents powered by large language models can reason and act using available tools to access emails, manage calendars, and push code to remo

When Does In-Context Search Help? A Sampling-Complexity Theory of Reflection-Driven Reasoning

Local AiDGX agent

arXiv:2607.06720v1 Announce Type: new Abstract: Training large language models (LLMs) with extended reasoning has enabled in-context search, in which models iteratively generate, critique, and revise

When Prompts Ignore Structure: Graph-Based Attribute Reasoning for Calibrated VLMs

ResearchDGX agent

arXiv:2607.07395v1 Announce Type: cross Abstract: Reliable confidence estimation remains a key limitation of test-time adaptation in vision-language models (VLMs), where prompt tuning improves zero-sh

Where Did the Variability Go? From Vibe Coding to Product Lines by Regeneration

ResearchDGX agent

arXiv:2606.19042v2 Announce Type: replace-cross Abstract: In vibe coding, an emerging AI-driven paradigm, an LLM generates an entire program from a natural language prompt, but what happens to the var

WHERE to Generate Matters: Budget-Aware Synthetic Augmentation for Label Skewed Federated Learning

SafetyDGX agent

arXiv:2607.06616v1 Announce Type: cross Abstract: Label skew in federated learning (FL) causes client drift and degrades global accuracy. Synthetic data augmentation can reduce this imbalance; however

Where to Intervene? Benchmarking Fairness-Aware Learning on Differentially Private Synthetic Tabular Data

Model ReleasesDGX agent

arXiv:2607.07471v1 Announce Type: cross Abstract: Machine learning models are increasingly deployed in high-stakes domains, raising concerns about both privacy and fairness. Differential Privacy (DP)

8 Jul 2026

A Definition and Roadmap for World Models

TutorialsDGX agent

arXiv:2607.06401v1 Announce Type: new Abstract: World models -- internal simulators that learn the structure and dynamics of an environment -- have become one of the most actively debated concepts in

A Guiding Framework for K-12 Teachers in Creating AI-powered Learning Technologies through Vibe Coding

ResearchDGX agent

arXiv:2607.05406v1 Announce Type: cross Abstract: Large language models generate code from natural language prompts, enabling 'vibe coding,' which allows non-programmers to develop computational solut

A Physics-Informed Neural Network Framework for Elastodynamic Wave Propagation in Bimaterial Systems

ResearchDGX agent

arXiv:2607.06479v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a promising framework for solving partial differential equations while embedding the underlying physica

A Three-Layer Framework for AI in Scientific Discovery

AgentsDGX agent

arXiv:2606.13566v2 Announce Type: replace Abstract: Current discussions of AI in scientific discovery are often dominated by two visible capabilities: search over existing knowledge and execution thro

A toy framework for single and multi-agent human-AI curiosity ecosystems

SafetyDGX agent

arXiv:2607.06214v1 Announce Type: new Abstract: This paper offers a toy framework for considering curiosity as an ecosystem. First, it suggests that a single agent's inquiry policy (how, when, and why

AbICL: In-Context Learning for Antigen-Specific Antibody Affinity Ranking

Model ReleasesDGX agent

arXiv:2607.05846v1 Announce Type: cross Abstract: Accurate ranking of antibody candidates according to their binding affinity is essential for therapeutic antibody discovery. However, existing methods

AdaStop: Cost-Aware Early Stopping for DNN Test Selection

ResearchDGX agent

arXiv:2607.05461v1 Announce Type: cross Abstract: Existing methods for testing deep neural networks (DNNs) primarily prioritize test inputs likely to reveal model faults under a fixed labeling budget.

Agentic AI for Commercial Insurance Underwriting with Adversarial Self-Critique

SafetyDGX agent

arXiv:2602.13213v2 Announce Type: replace Abstract: Commercial insurance underwriting is a labor-intensive process that requires manual review of extensive documentation to assess risk and determine p

Agentic AI for IPoDWDM Network Lifecycle Automation: An MCP-Enabled Architecture

AgentsDGX agent

arXiv:2607.05958v1 Announce Type: cross Abstract: We present a distributed, vendor-agnostic multi-MCP architecture for SDN-based automation and autonomous control of multi-vendor, multi-layer IPoDWDM

Agents That Teach: Towards Designing Incidental Learning Back into AI-Assisted Software Development

AgentsDGX agent

arXiv:2607.06101v1 Announce Type: cross Abstract: AI coding agents are rapidly reshaping how software is built, with developers increasingly delegating substantial coding tasks to autonomous agents in

AgoraSim: A Hybrid Agent-Based Modeling Framework

AgentsDGX agent

arXiv:2607.05999v1 Announce Type: new Abstract: LLM-agent simulations make natural-language social scenarios easy to instantiate, but their outputs can be overread as predictions and are often difficu

AI tools in Arab University English classrooms: Looking back and forward

ResearchDGX agent

arXiv:2607.05403v1 Announce Type: cross Abstract: This paper aims to synthesize empirical research on AI tools used to support English as a second/foreign language (EL2) learners in Arab University cl

aiAuthZ: Off-Host, Identity-Bound Authorization for AI Agents

Model ReleasesDGX agent

arXiv:2607.05518v1 Announce Type: cross Abstract: AI agents issue tool calls on the basis of text they cannot verify, so any party who controls part of the context can forge the appearance of authorit

AirflowAttack: Thermal-Airflow Adversarial Perturbations against Infrared Remote-Sensing Vision-Language Models

Model ReleasesDGX agent

arXiv:2607.06485v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed on infrared (IR) remote sensing imagery in security-critical settings, yet their adversarial r

Akashic: A Low-Overhead LLM Inference Service with MemAttention

AgentsDGX agent

arXiv:2607.05708v1 Announce Type: new Abstract: Recent LLM-based agent systems continuously accumulate context across multi-turn interactions, tool invocations, and cross-session workflows. Replaying

An Experimental Design Approach to Evaluating Agentic AI's Autonomous Model Discovery

Model ReleasesDGX agent

arXiv:2607.06413v1 Announce Type: cross Abstract: Large language model coding agents increasingly perform open-ended data modeling and analysis. These agents are stochastic and adaptive, and therefore

Analysis-by-Proxy: Localization Signals in VLMs Operating as Condition Encoders

Local AiDGX agent

arXiv:2607.06445v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly utilized as the conditioning backbone for diffusion-based image editing due to their remarkable multimo

ArtisanCAD: An Industrial-Level CAD Agent with Expert-Grounded Knowledge Distillation

Model ReleasesDGX agent

arXiv:2607.05750v1 Announce Type: new Abstract: Computer-aided design (CAD) for industrial components requires long-horizon procedural modeling, robust feature dependencies, editable parametric geomet

Auto-DSM Under the Lens: A Black-Box Evaluation Framework for LLM-Based DSM Generation

Model ReleasesDGX agent

arXiv:2607.05985v1 Announce Type: new Abstract: This paper presents a black-box evaluation framework to systematically assess the ability of Large Language Models (LLMs) to generate Design Structure M

Automated Recommendation of Programming Learning Content Using Pattern-based Knowledge Components

SafetyDGX agent

arXiv:2607.05409v1 Announce Type: cross Abstract: Introductory programming instruction relies on hands-on practice and short learning activities to support mastery of foundational concepts. Although m

BaFCo: A Document Understanding Benchmark for Complex Bangla Form Comprehension

Model ReleasesDGX agent

arXiv:2607.05614v1 Announce Type: cross Abstract: Document comprehension is a challenging yet impactful task for Multimodal Large Language Models, especially as these systems see growing adoption in r

Base Models Know How to Reason, Thinking Models Learn When

TutorialsDGX agent

arXiv:2510.07364v4 Announce Type: replace Abstract: What do thinking language models learn during training that their base models lack? We first present an unsupervised method that discovers a model's

Benchmarking KV-Cache Optimizations across Task Quality and System Performance for Long-Context Serving

Model ReleasesDGX agent

arXiv:2607.05399v1 Announce Type: cross Abstract: Large language model serving is increasingly limited by KV-cache growth under long-context workloads, yet existing KV-cache compression techniques are

Beyond Accuracy: How Humans Evaluate Legally Correct but Socially Controversial Legal Advice from Machines

ApplicationsDGX agent

arXiv:2607.05680v1 Announce Type: cross Abstract: AI systems are increasingly used to provide legal advice, raising questions about whether laypeople accept guidance from algorithms--especially when t

Beyond Correctness: Enhancing Architectural Reasoning in Code LLMs via Scalable Labeling with Agentic Judgment

AgentsDGX agent

arXiv:2606.14948v2 Announce Type: replace-cross Abstract: LLMs have substantially improved software engineering yet real-world development requires architectural understanding. Such understanding is p

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents

Model ReleasesDGX agent

arXiv:2510.19771v4 Announce Type: replace Abstract: LLM-based agents are increasingly moving towards proactivity: rather than awaiting instruction, they exercise agency to anticipate user needs and so

Beyond Refusal: A Same-Lineage Study of Aligned and Abliterated LLMs for Vulnerability Analysis

Model ReleasesDGX agent

arXiv:2607.05842v1 Announce Type: cross Abstract: Large language model (LLM)-assisted software security operates at a difficult boundary: the vulnerability-analysis terminology needed for legitimate c

Beyond Static Evaluation: Building Simulation Environments for Scalable Agentic Reinforcement Learning

AgentsDGX agent

arXiv:2607.05773v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, traditional static evaluation fails to capture multi-step decision-making. We introduce A

Beyond the Leaderboard: A Synthesis of Tool-Use, Planning, and Reasoning Failures in Large Language Model Agents

Model ReleasesDGX agent

arXiv:2607.05775v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly evaluated on their ability to use tools, plan multi-step tasks, coordinate with other agents, and ope

Binocular Gaze Estimation with Single Camera and Single Light Source

ResearchDGX agent

arXiv:2607.05473v1 Announce Type: cross Abstract: According to commonly consented theories, the minimum hardware requirement for gaze tracker is one camera and two light sources to realize gaze estima

Breaking Structural Isolation: Scalable Graph Clustering via Community-Aware Sampling and Structural Entropy

Model ReleasesDGX agent

arXiv:2607.05469v1 Announce Type: cross Abstract: Unsupervised graph clustering is a fundamental technique for uncovering underlying semantic patterns in large-scale networks. Although Graph Contrasti

Bridging Physical Reasoning and Task Generalization via Visual Action Outcome Reasoning Alignment

SafetyDGX agent

arXiv:2607.06522v1 Announce Type: new Abstract: Vision-language models (VLMs) struggle to generalize in interactive physical reasoning, particularly under unseen tasks and environments. Two key failur

CANONIC: Governance Is Compilation

Model ReleasesDGX agent

arXiv:2607.05410v1 Announce Type: cross Abstract: We present CANONIC: governed intelligence that compiles digital artifacts into an evidence ledger at scale. Large language models generate prose faste

CanvasAgent: Enabling Complex Image Creation and Editing via Visual Tool Orchestration

AgentsDGX agent

arXiv:2607.05465v1 Announce Type: cross Abstract: Complex image creation and editing often require more than a single generation or editing model. A user request may involve synthesizing images, local

Catalyst Papers in Artificial Intelligence Research: A Landscape on ICLR from 2017 to 2025

ResearchDGX agent

arXiv:2607.05401v1 Announce Type: cross Abstract: A small number of methodological contributions, including word2vec, the Transformer, large-scale pre-training, and reinforcement learning from human f

CCBENCH: Assessing LLM Cultural Competence via Implicitly Signaled Norms using Health Queries

ApplicationsDGX agent

arXiv:2607.05405v1 Announce Type: cross Abstract: To interact with users fairly and without stereotyping, AI models must display cultural competency, i.e., the ability to infer and adapt to a user's i

CHARLIE: An On-Premise Multi-Agent Retrieval-Augmented Generation System for Evidential Reasoning in Forensic Science

Local AiDGX agent

arXiv:2607.05428v1 Announce Type: cross Abstract: We present Charlie, an on-premise multi-agent Retrieval-Augmented Generation (RAG) system for structured evidential processing in digital forensic env

CMDR: Contextual Multimodal Document Retrieval

Model ReleasesDGX agent

arXiv:2607.05927v1 Announce Type: cross Abstract: Multimodal document retrieval aims to retrieve relevant pages while preserving both textual and visual content from the original document. However, ex

Complementary Roles of Image Classification and Vessel Segmentation in AI-Based Screening for Retinopathy of Prematurity Plus Disease in a Kenyan Preterm Cohort

ResearchDGX agent

arXiv:2607.05825v1 Announce Type: cross Abstract: Background. Retinopathy of prematurity (ROP) is a preventable cause of childhood blindness, with rising burden in low- and middle-income countries whe

Contrastive Predictive Coding with Compression for Enhanced Channel State Feedback in Wireless Networks

ResearchDGX agent

arXiv:2607.05419v1 Announce Type: cross Abstract: Accurate and timely channel state information (CSI) is essential for next-generation wireless systems, yet existing works treat CSI compression and CS

Controlling Tool Use with Heading-Specific Activation Steering

SafetyDGX agent

arXiv:2607.05790v1 Announce Type: new Abstract: Tool-augmented large language models extend their capabilities beyond parametric knowledge through external tools, but tend to invoke them unnecessarily

CSTutorBench: Benchmarking Small Language Models as Tutors for Block-Based Programming

Model ReleasesDGX agent

arXiv:2607.05571v1 Announce Type: new Abstract: Large language models are increasingly explored as AI tutors, yet deploying them in K-12 settings raises concerns around privacy, cost, and reliance on

Danus: Orchestrating Mathematical Reasoning Agents with Fact-Graph Memory

AgentsDGX agent

arXiv:2607.06447v1 Announce Type: new Abstract: Recent LLM-based mathematical reasoning agents have begun to tackle research-level problems and, in several cases, have contributed to the resolution of

Data Analysis in the Wild: Benchmarking Large Language Models Against Real-World Data Complexities

Model ReleasesDGX agent

arXiv:2607.06482v1 Announce Type: cross Abstract: Current benchmarks for evaluating Large Language Models (LLMs) in data analysis often fail to reflect real-world settings. They typically focus on fac

Data-dependent Evaluations for Budgeted Submodular Maximization

ApplicationsDGX agent

arXiv:2607.05759v1 Announce Type: cross Abstract: Submodular maximization is an important building block for developing algorithms in many areas such as machine learning and data mining. Due to the NP

Decision-Focused Scenario Generation and Selection for Efficient and Robust Grid Dispatch

ResearchDGX agent

arXiv:2607.05830v1 Announce Type: cross Abstract: The increasing uncertainty from flexible demand and renewable generation has made distributionally robust optimization (DRO) an important tool for rob

← Previous
1…7980818283…358
Next →