AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Automated conjecturing with TxGraffiti

DGX agent

arXiv:2409.19379v2 Announce Type: replace-cross Abstract: TxGraffiti is a data-driven, heuristic-based computer program developed to automate the process of generating conjectures across various mathe

researcharxiv-cs-ai
12 May 2026
Safety

Autonomous FAIR Digital Objects: From Passive Assertions to Active Knowledge

DGX agent
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2605.10370v1 Announce Type: new Abstract: Scientific knowledge on the Web is published as passive assertions and cannot decide when to validate evidence, reconcile contradictions, or update conf

safetyarxiv-cs-ai
12 May 2026
Research

BaLoRA: Bayesian Low-Rank Adaptation of Large Scale Models

DGX agent

arXiv:2605.08110v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the standard for fine-tuning large pre-trained models at reduced computational cost. However, its low-rank point

researcharxiv-cs-ai
12 May 2026
Research

Bangla-WhisperDiar: Fine-Tuning Whisper and PyAnnote for Bangla Long-Form Speech Recognition and Speaker Diarization

DGX agent

arXiv:2605.08214v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) and speaker diarization in Bangla remain challenging due to long form recordings, diverse acoustic conditions, and

researcharxiv-cs-ai
12 May 2026
Research

Batch Bayesian Active Learning with Partial Batch Label Sampling

DGX agent

arXiv:2510.09877v3 Announce Type: replace-cross Abstract: Over the past couple of decades, many active learning acquisition functions have been proposed, leaving practitioners with an unclear choice o

researcharxiv-cs-ai
12 May 2026
Agents

Batch-of-Thought: Cross-Instance Learning for Enhanced LLM Reasoning

DGX agent

arXiv:2601.02950v3 Announce Type: replace Abstract: Current Large Language Model reasoning systems process queries independently, discarding valuable cross-instance signals such as shared reasoning pa

agentsarxiv-cs-ai
12 May 2026
Model Releases

BEACON: A Multimodal Dataset for Learning Behavioral Fingerprints from Gameplay Data

DGX agent

arXiv:2605.10867v1 Announce Type: cross Abstract: Continuous authentication in high-stakes digital environments requires datasets with fine-grained behavioral signals under realistic cognitive and mot

model-releasesarxiv-cs-ai
12 May 2026
Safety

Behavioral Determinants of Deployed AI Agents in Social Networks: A Multi-Factor Study of Personality, Model, and Guardrail Specification

DGX agent

arXiv:2605.08463v1 Announce Type: new Abstract: Autonomous AI agents are increasingly deployed in open social environments, yet the relationship between their configuration specifications and their em

safetyarxiv-cs-ai
12 May 2026
Local Ai

Belief or Circuitry? Causal Evidence for In-Context Graph Learning

DGX agent

arXiv:2605.08405v1 Announce Type: new Abstract: How do LLMs learn in-context? Is it by pattern-matching recent tokens, or by inferring latent structure? We probe this question using a toy graph random

local-aiarxiv-cs-ai
12 May 2026
Model Releases

BenchCAD: A Comprehensive, Industry-Standard Benchmark for Programmatic CAD

DGX agent

arXiv:2605.10865v1 Announce Type: new Abstract: Industrial Computer-Aided Design (CAD) code generation requires models to produce executable parametric programs from visual or textual inputs. Beyond r

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Benchmarking Compositional Generalisation for Machine Learning Interatomic Potentials

DGX agent

arXiv:2605.08988v1 Announce Type: cross Abstract: Machine Learning Interatomic Potentials play a fundamental role in computational chemistry and materials science, enabling applications from molecular

model-releasesarxiv-cs-ai
12 May 2026
Research

Benchmarking ResNet Backbones in RT-DETR: Impact of Depth and Regularization under environmental conditions

DGX agent

arXiv:2605.08136v1 Announce Type: cross Abstract: Visual perception plays a central role in competitive robotics, where environmental variations can directly affect real-time detection performance. Th

researcharxiv-cs-ai
12 May 2026
Model Releases

Benchmarking Safety Risks of Knowledge-Intensive Reasoning under Malicious Knowledge Editing

DGX agent

arXiv:2605.10146v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on knowledge editing to support knowledge-intensive reasoning, but this flexibility also introduces criti

model-releasesarxiv-cs-ai
12 May 2026
Research

Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis

DGX agent

arXiv:2407.12173v2 Announce Type: replace-cross Abstract: Generative diffusion models have emerged as a powerful tool for high-quality image synthesis, yet their iterative nature demands significant c

researcharxiv-cs-ai
12 May 2026
Model Releases

Beyond Accuracy: Evaluating Strategy Diversity in LLM Mathematical Reasoning

DGX agent

arXiv:2605.09292v1 Announce Type: new Abstract: Large language models now achieve high final-answer accuracy on mathematical reasoning benchmarks, but accuracy alone does not capture reasoning flexibi

model-releasesarxiv-cs-ai
12 May 2026
Safety

Beyond Autonomy: A Dynamic Tiered AgentRunner Framework for Governable and Resilient Enterprise AI Execution

DGX agent

arXiv:2605.10223v1 Announce Type: new Abstract: Current large language model agent frameworks prioritize autonomy but lack the governability mechanisms required for enterprise deployment. High-risk wr

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs

DGX agent

arXiv:2605.09268v1 Announce Type: cross Abstract: Users interacting with Large Language Models (LLMs) in a multi-turn conversation routinely refine their requests or pivot to new topics. LLMs, however

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond ESG Scores: Learning Dynamic Constraints for Sequential Portfolio Optimization

DGX agent

arXiv:2605.09310v1 Announce Type: new Abstract: ESG-aware portfolio optimization is increasingly important for sustainable capital allocation, yet most learning-based methods still operationalize ESG

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning

DGX agent

arXiv:2605.08202v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) faces a critical challenge of overestimating the value of out-of-distribution (OOD) actions. Existing methods miti

safetyarxiv-cs-ai
12 May 2026
Safety

Beyond Self-Play: Hierarchical Reasoning for Continuous Motion in Closed-Loop Traffic Simulation

DGX agent

arXiv:2605.09153v1 Announce Type: cross Abstract: Closed-loop traffic simulation requires agents that are both scalable and behaviorally realistic. Recent self-play reinforcement learning approaches d

safetyarxiv-cs-ai
12 May 2026
Model Releases

Beyond the False Trade-off: Adaptive EWC for Stealthy and Generalizable T2I Backdoors

DGX agent

arXiv:2605.08280v1 Announce Type: cross Abstract: Preserving model fidelity is essential for stealthy text-to-image (T2I) backdoor attacks. Existing methods such as Learning without Forgetting (LwF) r

model-releasesarxiv-cs-ai
12 May 2026
Research

Beyond the Last Layer: Multi-Layer Representation Fusion for Visual Tokenizatio

DGX agent

arXiv:2605.10780v1 Announce Type: cross Abstract: Representation autoencoders that reuse frozen pretrained vision encoders as visual tokenizers have achieved strong reconstruction and generation quali

researcharxiv-cs-ai
12 May 2026
Model Releases

Beyond the Singular: Revealing the Value of Multiple Generations in Benchmark Evaluation

DGX agent

arXiv:2502.08943v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated significant utility in real-world applications, exhibiting impressive capabilities in natural l

model-releasesarxiv-cs-ai
12 May 2026
Safety

Bias by Necessity: Impossibility Theorems for Sequential Processing with Convergent AI and Human Validation

DGX agent

arXiv:2605.08716v1 Announce Type: new Abstract: Are certain cognitive biases mathematically inevitable consequences of sequential information processing? We prove that primacy effects, anchoring, and

safetyarxiv-cs-ai
12 May 2026
Safety

Big AI is accelerating the metacrisis: What can we do?

DGX agent

arXiv:2512.24863v2 Announce Type: replace-cross Abstract: The world is in the grip of ecological, meaning, and language crises that are converging into a metacrisis. Big AI is accelerating them all. L

safetyarxiv-cs-ai
12 May 2026
Safety

Biological Plausibility and Representational Alignment of Feedback Alignment in Convolutional Networks

DGX agent

arXiv:2605.08564v1 Announce Type: new Abstract: The feedback alignment (FA) algorithm offers a biologically plausible alternative to backpropagation (BP) for training neural networks yet notably fails

safetyarxiv-cs-ai
12 May 2026
Applications

Biosignal Fingerprinting: A Cross-Modal PPG-ECG Foundation Model

DGX agent

arXiv:2605.09579v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of global mortality, yet scalable cardiac monitoring is hindered by the gap between diagnostic-rich E

applicationsarxiv-cs-ai
12 May 2026
Safety

BONSAI: Bayesian Optimization with Natural Simplicity and Interpretability

DGX agent

arXiv:2602.07144v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) is a popular technique for sample-efficient optimization of black-box functions. In many applications, the paramete

safetyarxiv-cs-ai
12 May 2026
Research

BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models

DGX agent

arXiv:2605.09134v1 Announce Type: new Abstract: Reinforcement learning for program repair is hindered by sparse execution feedback and coarse sequence-level rewards that obscure which edits actually f

researcharxiv-cs-ai
12 May 2026
Safety

Break the Brake, Not the Wheel: Untargeted Jailbreak via Entropy Maximization

DGX agent

arXiv:2605.10764v1 Announce Type: cross Abstract: Recent studies show that gradient-based universal image jailbreaks on vision-language models (VLMs) exhibit little or no cross-model transferability,

safetyarxiv-cs-ai
12 May 2026
Research

Breaking Contextual Inertia: Reinforcement Learning with Single-Turn Anchors for Stable Multi-Turn Interaction

DGX agent

arXiv:2603.04783v2 Announce Type: replace Abstract: While LLMs demonstrate strong reasoning capabilities when provided with full information in a single turn, they exhibit substantial vulnerability in

researcharxiv-cs-ai
12 May 2026
Safety

Breaking the Grid: Distance-Guided Reinforcement Learning in Large Discrete Action Spaces

DGX agent

arXiv:2602.08616v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) is increasingly applied to large-scale decision-making problems like logistics, scheduling, and recommender system

safetyarxiv-cs-ai
12 May 2026
Model Releases

Bridging Modalities, Spanning Time: Structured Memory for Ultra-Long Agentic Video Reasoning

DGX agent

arXiv:2605.08271v1 Announce Type: cross Abstract: Understanding ultra-long videos such as egocentric recordings, live streams, or surveillance footage spanning days to weeks, remains a challenge. For

model-releasesarxiv-cs-ai
12 May 2026
Research

Bridging Sequence and Graph Structure for Epigenetic Age Prediction

DGX agent

arXiv:2605.10541v1 Announce Type: new Abstract: Epigenetic clocks based on DNA methylation have emerged as powerful tools for estimating biological age, with broad applications in aging research, age-

researcharxiv-cs-ai
12 May 2026
Agents

Bridging the Cognitive Gap: A Unified Memory Paradigm for 6G Agentic AI-RAN

DGX agent

arXiv:2605.10036v1 Announce Type: cross Abstract: As 6G evolves, the radio access network must transcend traditional automation to embrace agentic AI capable of perception, reasoning, and evolution. A

agentsarxiv-cs-ai
12 May 2026
Research

BubbleSpec: Turning Long-Tail Bubbles into Speculative Rollout Drafts for Synchronous Reinforcement Learning

DGX agent

arXiv:2605.08862v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has become a cornerstone for improving the performance of Large Language Models (LLMs). However, its rollout phase constit

researcharxiv-cs-ai
12 May 2026
Safety

Budget-Efficient Automatic Algorithm Design via Code Graph

DGX agent

arXiv:2605.10598v1 Announce Type: new Abstract: Large language models (LLMs) have emerged as powerful tools for automatic algorithm design (AAD). However, existing pipelines remain inefficient. They o

safetyarxiv-cs-ai
12 May 2026
Model Releases

Built Environment Reasoning from Remote Sensing Imagery Using Large Vision--Language Models

DGX agent

arXiv:2605.08404v1 Announce Type: cross Abstract: This work investigates the use of large language models (LLMs) for tasks in smart cities. The core idea is to leverage remote sensing imagery to chara

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

bViT: Investigating Single-Block Recurrence in Vision Transformers for Image Recognition

DGX agent

arXiv:2605.10661v1 Announce Type: cross Abstract: Vision Transformers (ViTs) are built by stacking independently parameterized blocks, but it remains unclear how much of this depth requires layer spec

model-releasesarxiv-cs-ai
12 May 2026
Local Ai

C2L-Net: A Data-Driven Model for State-of-Charge Estimation of Lithium-Ion Batteries During Discharge

DGX agent

arXiv:2605.08653v1 Announce Type: new Abstract: Accurate state-of-charge (SOC) estimation is critical for the safe and efficient operation of lithium-ion batteries in battery management systems (BMS).

local-aiarxiv-cs-ai
12 May 2026
Research

CachePrune: Teaching LLMs What Not to Follow via KV-Cache Editing

DGX agent

arXiv:2504.21228v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are susceptible to indirect prompt injection attacks, where the model inadvertently responds to instructions inje

researcharxiv-cs-ai
12 May 2026
Model Releases

CADBench: A Multimodal Benchmark for AI-Assisted CAD Program Generation

DGX agent

arXiv:2605.10873v1 Announce Type: cross Abstract: Recovering editable CAD programs from images or 3D observations is central to AI-assisted design, but progress is difficult to measure because existin

model-releasesarxiv-cs-ai
12 May 2026
Safety

CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs

DGX agent

arXiv:2605.09823v1 Announce Type: cross Abstract: We introduce CalBench, a controlled evaluation environment for studying multi-agent coordination through calendar scheduling. In CalBench, N agents ea

safetyarxiv-cs-ai
12 May 2026
Safety

CAMAL: Improving Attention Alignment and Faithfulness with Segmentation Masks

DGX agent

arXiv:2605.08325v1 Announce Type: cross Abstract: Many vision datasets now provide segmentation masks in addition to annotated images to support a wide range of tasks. In this work, we propose Class A

safetyarxiv-cs-ai
12 May 2026
Model Releases

Can Agent Benchmarks Support Their Scores? Evidence-Supported Bounds for Interactive-Agent Evaluation

DGX agent

arXiv:2605.10448v1 Announce Type: new Abstract: Interactive agent benchmarks map an agent run to a binary outcome through outcome checks. When these checks rely on surface level signals or fail to cap

model-releasesarxiv-cs-ai
12 May 2026
Research

Can Language Models Analyze Data? Evaluating Large Language Models for Question Answering over Datasets

DGX agent

arXiv:2605.10419v1 Announce Type: cross Abstract: This paper investigates the effectiveness of large language models (LLMs) in answering questions over datasets. We examine their performance in two sc

researcharxiv-cs-ai
12 May 2026
Safety

Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction

DGX agent

arXiv:2512.18880v2 Announce Type: replace-cross Abstract: Accurate estimation of item (question or task) difficulty is critical for educational assessment but suffers from the cold start problem. Whil

safetyarxiv-cs-ai
12 May 2026
Model Releases

Can LLMs Predict Polymer Physics Just by Reading Synthesis and Processing Prose?

DGX agent

arXiv:2605.08255v1 Announce Type: cross Abstract: Can large language models predict physical and mechanical polymer properties simply by reading unstructured scientific prose? Polymer performance is r

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…329330331332333…448
Next →