AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Applications

Using AI-based Learning Assistants in Higher Education: A Large-Scale Descriptive Analysis

DGX agent

arXiv:2607.08748v1 Announce Type: new Abstract: In this study, we present a large-scale descriptive analysis of the use of an AI-based learning assistant (Syntea) in higher education. Based on objecti

applicationsarxiv-cs-ai
10 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Validity of LLMs as data annotators: AMALIA on authority

DGX agent

arXiv:2607.08731v1 Announce Type: cross Abstract: A national language model offers a linguistic community its own instrument for measuring what its citizens say and value. Portugal's AMALIA, a publicl

model-releasesarxiv-cs-ai
10 Jul 2026
Tutorials

VectorizationLLM: Smart Vectorization Based AI Assistant

DGX agent

arXiv:2607.07846v1 Announce Type: new Abstract: VectorizationLLM is a specialized Large Language Model based on Google open-weight LLMs. The model is designed to assist students to learn smart vectori

tutorialsarxiv-cs-ai
10 Jul 2026
Research

VEGAS: Human-Aligned Video Caption Evaluation via Gaze

DGX agent

arXiv:2607.08489v1 Announce Type: cross Abstract: Vision-language models excel at video captioning, yet typically generate descriptions that fail to capture individual viewers' attention. We propose V

researcharxiv-cs-ai
10 Jul 2026
Research

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval

DGX agent

arXiv:2607.08541v1 Announce Type: cross Abstract: Open-vocabulary object detection and segmentation aim to recognize arbitrary objects beyond predefined categories. Although recent vision-language and

researcharxiv-cs-ai
10 Jul 2026
Model Releases

WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving

DGX agent

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition o

model-releasesarxiv-cs-ai
10 Jul 2026
Local Ai

WebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web Search

DGX agent

arXiv:2607.08662v1 Announce Type: cross Abstract: Large language model (LLM)-based web search agents are transforming information seeking from simple factoid question answering into complex, deep-and-

local-aiarxiv-cs-ai
10 Jul 2026
Research

What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration and Faithfulness

DGX agent

arXiv:2607.08046v1 Announce Type: cross Abstract: Large language models fine-tuned for forecasting can be accurate yet poorly calibrated, and their chain-of-thought (CoT) reasoning may not faithfully

researcharxiv-cs-ai
10 Jul 2026
Research

When Implausible Tokens Get Reinforced: Tail-Aware Credit Calibration for LLM Reinforcement Learning

DGX agent

arXiv:2607.07976v1 Announce Type: cross Abstract: Reinforcement learning (RL) has achieved remarkable success in enhancing the reasoning capabilities of large language models (LLMs). However, widely u

researcharxiv-cs-ai
10 Jul 2026
Model Releases

When LLMs Agree, Are They Right? Auditing Self-Consistency and Cross-Model Agreement as Confidence Signals

DGX agent

arXiv:2607.08065v1 Announce Type: new Abstract: LLM-as-judge (Zheng et al., 2023) is increasingly the default for evaluating AI systems in enterprise pipelines, often scaled to ensembles (Verga et al.

model-releasesarxiv-cs-ai
10 Jul 2026
Safety

When Structured Sparse Autoencoders Learn Consistent Concepts Across Modalities

DGX agent

arXiv:2607.08605v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have emerged as a promising technique for mechanistic interpretability by learning a set of sparse latent features in large

safetyarxiv-cs-ai
10 Jul 2026
Safety

When Synthetic Speech Is All You Have: Better Call GRPO

DGX agent

arXiv:2607.08409v1 Announce Type: cross Abstract: LLM-based ASR adapted to regulated domains such as banking is bottlenecked by privacy: real speech is costly and legally constrained to collect, makin

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

When the Judge Changes, So Does the Measurement: Auditing LLM-as-Judge Reliability

DGX agent

arXiv:2607.08535v1 Announce Type: cross Abstract: An LLM-as-judge score can move even when the candidate responses stay fixed, simply because the evaluator has changed. We treat this evaluator-replace

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

When Thinking Hurts: Epistemic Signals in the Reasoning Chains of Visual Language Models

DGX agent

arXiv:2607.08059v1 Announce Type: cross Abstract: Uncertainty quantification for visual language models (VLMs) conventionally targets the answer token distribution. We provide the first three-family e

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

Who Analyses the Analyser? Self-Validating LLM Hazard Analysis with Constitutional Meta-STPA

DGX agent

arXiv:2607.08054v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly trusted to draft the artifacts of safety analysis such as, losses, hazards, Unsafe Control Actions (UCAs

model-releasesarxiv-cs-ai
10 Jul 2026
Local Ai

Who Broke the System? Failure Localization in LLM-Based Multi-Agent Systems

DGX agent

arXiv:2607.07989v1 Announce Type: cross Abstract: Large language model (LLM) based multi-agent systems enable complex problem solving through coordinated reasoning and action, but their distributed st

local-aiarxiv-cs-ai
10 Jul 2026
Safety

Workflow as Knowledge: Semantic Persistence for LLM-Mediated Workflows

DGX agent

arXiv:2607.08740v1 Announce Type: new Abstract: Large language model (LLM) applications increasingly use explicit workflows for tool use, retrieval, branching, checkpointing, and human approval. Exist

safetyarxiv-cs-ai
10 Jul 2026
Safety

XFACTORS: Disentangled Information Bottleneck via Contrastive Supervision

DGX agent

arXiv:2601.21688v2 Announce Type: replace-cross Abstract: Disentangled representation learning aims to map independent factors of variation to independent representation components. On one hand, purel

safetyarxiv-cs-ai
10 Jul 2026
Tutorials

A Continual Learning Framework for Adaptive Control of Modular Soft Robots

DGX agent

arXiv:2607.06740v1 Announce Type: cross Abstract: Soft robots have attracted significant attention in applications such as medical intervention, rehabilitation, and robotic manipulation due to their i

tutorialsarxiv-cs-ai
9 Jul 2026
Model Releases

A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong

DGX agent

arXiv:2607.06854v1 Announce Type: cross Abstract: Reinforcement learning agents for imperfect-information card games are only as strong as the opponents they train against, and they are hard to grade,

model-releasesarxiv-cs-ai
9 Jul 2026
Research

A Multi-Analyst LLM Pipeline for Auditable Rule Discovery Across 68 Public Physiological Corpora

DGX agent

arXiv:2607.06802v1 Announce Type: cross Abstract: Open physiological corpora are heterogeneous: they use different sensors, labels, sampling rates, recording settings, and clinical endpoints. They can

researcharxiv-cs-ai
9 Jul 2026
Model Releases

A Study of Commonsense Reasoning over Visual Object Properties

DGX agent

arXiv:2508.10956v3 Announce Type: replace-cross Abstract: Inspired by human categorization, visual reasoning about object properties, such as physical attributes and functions, involves identifying an

model-releasesarxiv-cs-ai
9 Jul 2026
Research

Accurate, Interdisciplinary and Transparent Structure-property Understanding with Deep Native Structural Reasoning

DGX agent

arXiv:2607.07708v1 Announce Type: cross Abstract: Structure-property relationships are foundational to biology, chemistry and materials science, where function, reactivity and physical response emerge

researcharxiv-cs-ai
9 Jul 2026
Safety

Ad Headline Generation using Self-Critical Masked Language Model

DGX agent

arXiv:2607.06818v1 Announce Type: cross Abstract: For any E-commerce website it is a nontrivial problem to build enduring advertisements that attract shoppers. It is hard to pass the creative quality

safetyarxiv-cs-ai
9 Jul 2026
Safety

AGAPI-Agents: An Open-Access Agentic AI Platform for Accelerated Materials Design on AtomGPT.org

DGX agent

arXiv:2512.11935v2 Announce Type: replace Abstract: Agentic AI systems increasingly connect large language models (LLMs) to external scientific tools, yet whether and when tool access improves predict

safetyarxiv-cs-ai
9 Jul 2026
Safety

Agentic Data Environments

DGX agent

arXiv:2607.07397v1 Announce Type: new Abstract: Autonomous agents promise substantial gains in speed, scale, and labor efficiency, but their failures can impose abrupt and often irreversible costs. Th

safetyarxiv-cs-ai
9 Jul 2026
Model Releases

AgentLens: Production-Assessed Trajectory Reviews for Coding Agent Evaluation

DGX agent

arXiv:2607.06624v1 Announce Type: new Abstract: We present AgentLens, a production-assessed benchmark for interactive code agents. Most code-agent benchmarks reduce a run to a single bit -- did the ta

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Agon: Competitive Cross-Model RL with Implicit Rival Grading of Reasoning

DGX agent

arXiv:2607.07690v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards (e.g. GRPO) is the engine behind today's reasoning models, yet it grades only the final answer. On hard

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

AI Chatbot Suicide Risk Detection and Response: Human Validation Study of the Open-Source VERA-MH Safety Evaluation

DGX agent

arXiv:2602.05088v4 Announce Type: replace Abstract: Millions of people now use generative AI chatbots for psychological support. Despite their promise, the most pressing question in AI for mental heal

model-releasesarxiv-cs-ai
9 Jul 2026
Research

AirPASS: Over-the-Air Federated Learning via Pinching Antenna Systems

DGX agent

arXiv:2607.06768v1 Announce Type: cross Abstract: This paper investigates over-the-air federated learning (AirFL) in wireless systems where the access point is equipped with a multi-waveguide pinching

researcharxiv-cs-ai
9 Jul 2026
Local Ai

ALER-TI: Aligned Latent Embedding Retrieval for Time Series Imputation

DGX agent

arXiv:2607.07640v1 Announce Type: cross Abstract: Deep learning has significantly advanced time series imputation, yet most existing architectures primarily rely on localized temporal context within t

local-aiarxiv-cs-ai
9 Jul 2026
Tutorials

An Adaptive Differentially Private Federated Learning Framework

DGX agent

arXiv:2602.06838v3 Announce Type: replace Abstract: Federated learning enables collaborative model training across distributed clients while preserving data privacy. However, in practical deployments,

tutorialsarxiv-cs-ai
9 Jul 2026
Local Ai

AnchorPrune: Relevance-Anchored Contextual Expansion for Visual Token Pruning

DGX agent

arXiv:2607.07033v1 Announce Type: cross Abstract: Large vision-language models incur substantial inference costs because high-resolution inputs introduce thousands of visual tokens, many of which are

local-aiarxiv-cs-ai
9 Jul 2026
Applications

Anomaly detection in time-series via inductive biases in the latent space of conditional normalizing flows

DGX agent

arXiv:2603.11756v2 Announce Type: replace Abstract: Deep generative models for anomaly detection in multivariate time-series are typically trained by maximizing observed data likelihood. However, like

applicationsarxiv-cs-ai
9 Jul 2026
Research

AT-Attn: Temporal-Aware Cross-Attention for Longitudinal Multimodal Alzheimer's Disease Diagnosis

DGX agent

arXiv:2607.07091v1 Announce Type: cross Abstract: In longitudinal Alzheimer's disease (AD) diagnosis support, clinical and imaging information is often collected at irregular visits. Integrating these

researcharxiv-cs-ai
9 Jul 2026
Model Releases

At-Grok Is Not Converged:A Measurement-Validity Audit for Grokking Representation Metrics

DGX agent

arXiv:2607.06639v1 Announce Type: cross Abstract: On modular arithmetic, a network's embedding keeps compressing for tens of thousands of steps after it has already generalized. Reading effective rank

model-releasesarxiv-cs-ai
9 Jul 2026
Research

Audio Sentiment Analysis via Distillation and Cross-Modal Integration of Generated Multilingual Transcripts

DGX agent

arXiv:2607.06611v1 Announce Type: cross Abstract: Automatically recognizing the sentiment, positive or negative, from speech is a challenging task, requiring both the analysis of vocal inflections and

researcharxiv-cs-ai
9 Jul 2026
Tutorials

Bayesian Optimization of Genetic Algorithm Hyperparameters in a Multi-Fidelity Framework for Efficient Lattice Material Design

DGX agent

arXiv:2607.07289v1 Announce Type: cross Abstract: This study presents a multi-fidelity framework for the systematic optimization of genetic algorithm (GA) hyperparameters. The framework integrates thr

tutorialsarxiv-cs-ai
9 Jul 2026
Safety

Behavior Foundations for Quadruped Robots: ABot-C0 Technical Report

DGX agent

arXiv:2607.07370v1 Announce Type: cross Abstract: In embodied intelligence systems, the motion controller serves as the critical bridge between semantic reasoning and physical execution. Humanoid cont

safetyarxiv-cs-ai
9 Jul 2026
Model Releases

Beyond Attack-Success Rate: Action-Graded Severity Scale for Tool-Using AI Agents

DGX agent

arXiv:2607.07474v1 Announce Type: cross Abstract: Agentic red-teaming benchmarks report whether an injected agent was compromised as a single bit: the attack succeeded, or it did not. We argue that th

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Breaking Database Lock-in: Agentic Regeneration of High Performance Storage Readers for Database Bypass

DGX agent

arXiv:2607.07696v1 Announce Type: cross Abstract: Analytical workloads operating on data stored in external database systems face a fundamental bottleneck: data access is guarded entirely by the datab

model-releasesarxiv-cs-ai
9 Jul 2026
Research

ButterflyMoE: Compression-Scalable Ternary Experts via Structured Butterfly Orbits

DGX agent

arXiv:2601.13563v5 Announce Type: replace-cross Abstract: In current Mixture of Experts (MoE) architectures, linear memory scaling is present, the memory grows as the number of experts increases. N in

researcharxiv-cs-ai
9 Jul 2026
Model Releases

Can Reinforcement Learning Efficiently Discover Price Manipulation?

DGX agent

arXiv:2607.06121v1 Announce Type: cross Abstract: In this paper, we investigate whether a model-free RL agent can identify and exploit price manipulation opportunities more effectively than a traditio

model-releasesarxiv-cs-ai
9 Jul 2026
Tutorials

Can We Really Learn One Representation to Optimize All Rewards?

DGX agent

arXiv:2602.11399v2 Announce Type: replace-cross Abstract: As unsupervised pretraining becomes increasingly ubiquitous in reinforcement learning, a more thorough theoretical understanding of these meth

tutorialsarxiv-cs-ai
9 Jul 2026
Research

CarbonCLIP: Enhance Carbon Prediction from Satellite Imagery via Integrated Street-View Semantics and Temporal Context Training

DGX agent

arXiv:2607.07292v1 Announce Type: cross Abstract: Accurately estimating urban carbon emissions is critical for sustainable urban planning, yet many existing approaches remain difficult to apply consis

researcharxiv-cs-ai
9 Jul 2026
Safety

CARLA-GS: Decoupling Representation, Reasoning, and Physics Simulation for Autonomous Driving Corner-Case Synthesis

DGX agent

arXiv:2607.07601v1 Announce Type: cross Abstract: Safety evaluation for autonomous driving is dominated by rare, safety-critical interactions, motivating simulators that can deliberately synthesize co

safetyarxiv-cs-ai
9 Jul 2026
Model Releases

Co-LMLM: Continuous-Query Limited Memory Language Models

DGX agent

arXiv:2607.07707v1 Announce Type: cross Abstract: Limited memory language models (LMLMs) externalize factual knowledge during pretraining to a knowledge base (KB), rather than memorizing it in their w

model-releasesarxiv-cs-ai
9 Jul 2026
Research

Collaborative Synthetic Data Generation for Knowledge Transfer in Federated Learning

DGX agent

arXiv:2607.07565v1 Announce Type: cross Abstract: One-shot federated learning (OSFL) addresses the communication overhead of federated learning by limiting training to a single round, but doing so wit

researcharxiv-cs-ai
9 Jul 2026
← Previous
1…9596979899…448
Next →