AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Understanding Two-Layer Neural Networks with Smooth Activation Functions

DGX agent

arXiv:2507.14177v2 Announce Type: replace-cross Abstract: This paper aims to understand the training solution, which is obtained by the back-propagation algorithm, of two-layer neural networks whose h

researcharxiv-cs-ai
9 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Validate the Dream Before You Trust Its Verdict: Admissibility for World-Model Simulators

DGX agent

arXiv:2607.07196v1 Announce Type: cross Abstract: Across robotics, World Models (WMs) are increasingly used to evaluate action policies by simulating the consequences of actions in an imagined world,

safetyarxiv-cs-ai
9 Jul 2026
Safety

Vision Foundation Models in Radiology: A Scoping Review of Data, Methodology, Evaluation and Clinical Translation

DGX agent

arXiv:2607.07219v1 Announce Type: cross Abstract: Vision foundation models (VFMs) are increasingly being developed for radiological imaging, yet their definition, development and evaluation remain het

safetyarxiv-cs-ai
9 Jul 2026
Research

Vision Language Action (VLA) Models for Unmanned Aerial Robotics and Bimanual Manipulation: A Review

DGX agent

arXiv:2607.06706v1 Announce Type: cross Abstract: Vision Language Action (VLA) models unify visual perception, natural-language understanding, and action generation within a single foundation model, a

researcharxiv-cs-ai
9 Jul 2026
Research

VOTE: Vision-Language-Action Optimization with Trajectory Ensemble Voting

DGX agent

arXiv:2507.05116v5 Announce Type: replace-cross Abstract: Recent large-scale Vision Language Action (VLA) models have shown superior performance in robotic manipulation tasks guided by natural languag

researcharxiv-cs-ai
9 Jul 2026
Research

WAM-TTT: Steering World-Action Models by Watching Human Play at Test Time

DGX agent

arXiv:2607.06988v1 Announce Type: cross Abstract: Steering robot foundation models (RFMs) toward new task variants or user-preferred behaviors remains challenging, often requiring additional robot dem

researcharxiv-cs-ai
9 Jul 2026
Model Releases

What Predicts Correctness in Text-to-SQL? A Selective-Prediction Study

DGX agent

arXiv:2607.06799v1 Announce Type: cross Abstract: Evaluating uncertainty in AI-generated SQL queries requires estimating whether a query is correct, where correct means it executes to the same result

model-releasesarxiv-cs-ai
9 Jul 2026
Local Ai

When Agents Go Rogue: Activation-Based Detection of Malicious Behaviors in Multi-Agent Systems

DGX agent

arXiv:2607.06807v1 Announce Type: cross Abstract: While enabling effective collaboration on complex tasks, LLM-based Multi-Agent Systems (MAS) face critical security challenges due to vulnerabilities

local-aiarxiv-cs-ai
9 Jul 2026
Safety

When Agents Remember Too Much: Memory Poisoning Attacks on Large Language Model Agents

DGX agent

arXiv:2607.06595v1 Announce Type: cross Abstract: Personal AI agents powered by large language models can reason and act using available tools to access emails, manage calendars, and push code to remo

safetyarxiv-cs-ai
9 Jul 2026
Local Ai

When Does In-Context Search Help? A Sampling-Complexity Theory of Reflection-Driven Reasoning

DGX agent

arXiv:2607.06720v1 Announce Type: new Abstract: Training large language models (LLMs) with extended reasoning has enabled in-context search, in which models iteratively generate, critique, and revise

local-aiarxiv-cs-ai
9 Jul 2026
Research

When Prompts Ignore Structure: Graph-Based Attribute Reasoning for Calibrated VLMs

DGX agent

arXiv:2607.07395v1 Announce Type: cross Abstract: Reliable confidence estimation remains a key limitation of test-time adaptation in vision-language models (VLMs), where prompt tuning improves zero-sh

researcharxiv-cs-ai
9 Jul 2026
Research

Where Did the Variability Go? From Vibe Coding to Product Lines by Regeneration

DGX agent

arXiv:2606.19042v2 Announce Type: replace-cross Abstract: In vibe coding, an emerging AI-driven paradigm, an LLM generates an entire program from a natural language prompt, but what happens to the var

researcharxiv-cs-ai
9 Jul 2026
Safety

WHERE to Generate Matters: Budget-Aware Synthetic Augmentation for Label Skewed Federated Learning

DGX agent

arXiv:2607.06616v1 Announce Type: cross Abstract: Label skew in federated learning (FL) causes client drift and degrades global accuracy. Synthetic data augmentation can reduce this imbalance; however

safetyarxiv-cs-ai
9 Jul 2026
Model Releases

Where to Intervene? Benchmarking Fairness-Aware Learning on Differentially Private Synthetic Tabular Data

DGX agent

arXiv:2607.07471v1 Announce Type: cross Abstract: Machine learning models are increasingly deployed in high-stakes domains, raising concerns about both privacy and fairness. Differential Privacy (DP)

model-releasesarxiv-cs-ai
9 Jul 2026
Tutorials

A Definition and Roadmap for World Models

DGX agent

arXiv:2607.06401v1 Announce Type: new Abstract: World models -- internal simulators that learn the structure and dynamics of an environment -- have become one of the most actively debated concepts in

tutorialsarxiv-cs-ai
8 Jul 2026
Research

A Guiding Framework for K-12 Teachers in Creating AI-powered Learning Technologies through Vibe Coding

DGX agent

arXiv:2607.05406v1 Announce Type: cross Abstract: Large language models generate code from natural language prompts, enabling 'vibe coding,' which allows non-programmers to develop computational solut

researcharxiv-cs-ai
8 Jul 2026
Research

A Physics-Informed Neural Network Framework for Elastodynamic Wave Propagation in Bimaterial Systems

DGX agent

arXiv:2607.06479v1 Announce Type: new Abstract: Physics-informed neural networks (PINNs) provide a promising framework for solving partial differential equations while embedding the underlying physica

researcharxiv-cs-ai
8 Jul 2026
Agents

A Three-Layer Framework for AI in Scientific Discovery

DGX agent

arXiv:2606.13566v2 Announce Type: replace Abstract: Current discussions of AI in scientific discovery are often dominated by two visible capabilities: search over existing knowledge and execution thro

agentsarxiv-cs-ai
8 Jul 2026
Safety

A toy framework for single and multi-agent human-AI curiosity ecosystems

DGX agent

arXiv:2607.06214v1 Announce Type: new Abstract: This paper offers a toy framework for considering curiosity as an ecosystem. First, it suggests that a single agent's inquiry policy (how, when, and why

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

AbICL: In-Context Learning for Antigen-Specific Antibody Affinity Ranking

DGX agent

arXiv:2607.05846v1 Announce Type: cross Abstract: Accurate ranking of antibody candidates according to their binding affinity is essential for therapeutic antibody discovery. However, existing methods

model-releasesarxiv-cs-ai
8 Jul 2026
Research

AdaStop: Cost-Aware Early Stopping for DNN Test Selection

DGX agent

arXiv:2607.05461v1 Announce Type: cross Abstract: Existing methods for testing deep neural networks (DNNs) primarily prioritize test inputs likely to reveal model faults under a fixed labeling budget.

researcharxiv-cs-ai
8 Jul 2026
Safety

Agentic AI for Commercial Insurance Underwriting with Adversarial Self-Critique

DGX agent

arXiv:2602.13213v2 Announce Type: replace Abstract: Commercial insurance underwriting is a labor-intensive process that requires manual review of extensive documentation to assess risk and determine p

safetyarxiv-cs-ai
8 Jul 2026
Agents

Agentic AI for IPoDWDM Network Lifecycle Automation: An MCP-Enabled Architecture

DGX agent

arXiv:2607.05958v1 Announce Type: cross Abstract: We present a distributed, vendor-agnostic multi-MCP architecture for SDN-based automation and autonomous control of multi-vendor, multi-layer IPoDWDM

agentsarxiv-cs-ai
8 Jul 2026
Agents

Agents That Teach: Towards Designing Incidental Learning Back into AI-Assisted Software Development

DGX agent

arXiv:2607.06101v1 Announce Type: cross Abstract: AI coding agents are rapidly reshaping how software is built, with developers increasingly delegating substantial coding tasks to autonomous agents in

agentsarxiv-cs-ai
8 Jul 2026
Agents

AgoraSim: A Hybrid Agent-Based Modeling Framework

DGX agent

arXiv:2607.05999v1 Announce Type: new Abstract: LLM-agent simulations make natural-language social scenarios easy to instantiate, but their outputs can be overread as predictions and are often difficu

agentsarxiv-cs-ai
8 Jul 2026
Research

AI tools in Arab University English classrooms: Looking back and forward

DGX agent

arXiv:2607.05403v1 Announce Type: cross Abstract: This paper aims to synthesize empirical research on AI tools used to support English as a second/foreign language (EL2) learners in Arab University cl

researcharxiv-cs-ai
8 Jul 2026
Model Releases

aiAuthZ: Off-Host, Identity-Bound Authorization for AI Agents

DGX agent

arXiv:2607.05518v1 Announce Type: cross Abstract: AI agents issue tool calls on the basis of text they cannot verify, so any party who controls part of the context can forge the appearance of authorit

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

AirflowAttack: Thermal-Airflow Adversarial Perturbations against Infrared Remote-Sensing Vision-Language Models

DGX agent

arXiv:2607.06485v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed on infrared (IR) remote sensing imagery in security-critical settings, yet their adversarial r

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

Akashic: A Low-Overhead LLM Inference Service with MemAttention

DGX agent

arXiv:2607.05708v1 Announce Type: new Abstract: Recent LLM-based agent systems continuously accumulate context across multi-turn interactions, tool invocations, and cross-session workflows. Replaying

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

An Experimental Design Approach to Evaluating Agentic AI's Autonomous Model Discovery

DGX agent

arXiv:2607.06413v1 Announce Type: cross Abstract: Large language model coding agents increasingly perform open-ended data modeling and analysis. These agents are stochastic and adaptive, and therefore

model-releasesarxiv-cs-ai
8 Jul 2026
Local Ai

Analysis-by-Proxy: Localization Signals in VLMs Operating as Condition Encoders

DGX agent

arXiv:2607.06445v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly utilized as the conditioning backbone for diffusion-based image editing due to their remarkable multimo

local-aiarxiv-cs-ai
8 Jul 2026
Model Releases

ArtisanCAD: An Industrial-Level CAD Agent with Expert-Grounded Knowledge Distillation

DGX agent

arXiv:2607.05750v1 Announce Type: new Abstract: Computer-aided design (CAD) for industrial components requires long-horizon procedural modeling, robust feature dependencies, editable parametric geomet

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Auto-DSM Under the Lens: A Black-Box Evaluation Framework for LLM-Based DSM Generation

DGX agent

arXiv:2607.05985v1 Announce Type: new Abstract: This paper presents a black-box evaluation framework to systematically assess the ability of Large Language Models (LLMs) to generate Design Structure M

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

Automated Recommendation of Programming Learning Content Using Pattern-based Knowledge Components

DGX agent

arXiv:2607.05409v1 Announce Type: cross Abstract: Introductory programming instruction relies on hands-on practice and short learning activities to support mastery of foundational concepts. Although m

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

BaFCo: A Document Understanding Benchmark for Complex Bangla Form Comprehension

DGX agent

arXiv:2607.05614v1 Announce Type: cross Abstract: Document comprehension is a challenging yet impactful task for Multimodal Large Language Models, especially as these systems see growing adoption in r

model-releasesarxiv-cs-ai
8 Jul 2026
Tutorials

Base Models Know How to Reason, Thinking Models Learn When

DGX agent

arXiv:2510.07364v4 Announce Type: replace Abstract: What do thinking language models learn during training that their base models lack? We first present an unsupervised method that discovers a model's

tutorialsarxiv-cs-ai
8 Jul 2026
Model Releases

Benchmarking KV-Cache Optimizations across Task Quality and System Performance for Long-Context Serving

DGX agent

arXiv:2607.05399v1 Announce Type: cross Abstract: Large language model serving is increasingly limited by KV-cache growth under long-context workloads, yet existing KV-cache compression techniques are

model-releasesarxiv-cs-ai
8 Jul 2026
Applications

Beyond Accuracy: How Humans Evaluate Legally Correct but Socially Controversial Legal Advice from Machines

DGX agent

arXiv:2607.05680v1 Announce Type: cross Abstract: AI systems are increasingly used to provide legal advice, raising questions about whether laypeople accept guidance from algorithms--especially when t

applicationsarxiv-cs-ai
8 Jul 2026
Agents

Beyond Correctness: Enhancing Architectural Reasoning in Code LLMs via Scalable Labeling with Agentic Judgment

DGX agent

arXiv:2606.14948v2 Announce Type: replace-cross Abstract: LLMs have substantially improved software engineering yet real-world development requires architectural understanding. Such understanding is p

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents

DGX agent

arXiv:2510.19771v4 Announce Type: replace Abstract: LLM-based agents are increasingly moving towards proactivity: rather than awaiting instruction, they exercise agency to anticipate user needs and so

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Beyond Refusal: A Same-Lineage Study of Aligned and Abliterated LLMs for Vulnerability Analysis

DGX agent

arXiv:2607.05842v1 Announce Type: cross Abstract: Large language model (LLM)-assisted software security operates at a difficult boundary: the vulnerability-analysis terminology needed for legitimate c

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

Beyond Static Evaluation: Building Simulation Environments for Scalable Agentic Reinforcement Learning

DGX agent

arXiv:2607.05773v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into autonomous agents, traditional static evaluation fails to capture multi-step decision-making. We introduce A

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

Beyond the Leaderboard: A Synthesis of Tool-Use, Planning, and Reasoning Failures in Large Language Model Agents

DGX agent

arXiv:2607.05775v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly evaluated on their ability to use tools, plan multi-step tasks, coordinate with other agents, and ope

model-releasesarxiv-cs-ai
8 Jul 2026
Research

Binocular Gaze Estimation with Single Camera and Single Light Source

DGX agent

arXiv:2607.05473v1 Announce Type: cross Abstract: According to commonly consented theories, the minimum hardware requirement for gaze tracker is one camera and two light sources to realize gaze estima

researcharxiv-cs-ai
8 Jul 2026
Model Releases

Breaking Structural Isolation: Scalable Graph Clustering via Community-Aware Sampling and Structural Entropy

DGX agent

arXiv:2607.05469v1 Announce Type: cross Abstract: Unsupervised graph clustering is a fundamental technique for uncovering underlying semantic patterns in large-scale networks. Although Graph Contrasti

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

Bridging Physical Reasoning and Task Generalization via Visual Action Outcome Reasoning Alignment

DGX agent

arXiv:2607.06522v1 Announce Type: new Abstract: Vision-language models (VLMs) struggle to generalize in interactive physical reasoning, particularly under unseen tasks and environments. Two key failur

safetyarxiv-cs-ai
8 Jul 2026
Model Releases

CANONIC: Governance Is Compilation

DGX agent

arXiv:2607.05410v1 Announce Type: cross Abstract: We present CANONIC: governed intelligence that compiles digital artifacts into an evidence ledger at scale. Large language models generate prose faste

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

CanvasAgent: Enabling Complex Image Creation and Editing via Visual Tool Orchestration

DGX agent

arXiv:2607.05465v1 Announce Type: cross Abstract: Complex image creation and editing often require more than a single generation or editing model. A user request may involve synthesizing images, local

agentsarxiv-cs-ai
8 Jul 2026
← Previous
1…99100101102103…448
Next →