AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
12 May 2026

Automated conjecturing with TxGraffiti

ResearchDGX agent

arXiv:2409.19379v2 Announce Type: replace-cross Abstract: TxGraffiti is a data-driven, heuristic-based computer program developed to automate the process of generating conjectures across various mathe

Autonomous FAIR Digital Objects: From Passive Assertions to Active Knowledge

SafetyDGX agent

arXiv:2605.10370v1 Announce Type: new Abstract: Scientific knowledge on the Web is published as passive assertions and cannot decide when to validate evidence, reconcile contradictions, or update conf

BaLoRA: Bayesian Low-Rank Adaptation of Large Scale Models

ResearchDGX agent

arXiv:2605.08110v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) has become the standard for fine-tuning large pre-trained models at reduced computational cost. However, its low-rank point


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Bangla-WhisperDiar: Fine-Tuning Whisper and PyAnnote for Bangla Long-Form Speech Recognition and Speaker Diarization

ResearchDGX agent

arXiv:2605.08214v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) and speaker diarization in Bangla remain challenging due to long form recordings, diverse acoustic conditions, and

Batch Bayesian Active Learning with Partial Batch Label Sampling

ResearchDGX agent

arXiv:2510.09877v3 Announce Type: replace-cross Abstract: Over the past couple of decades, many active learning acquisition functions have been proposed, leaving practitioners with an unclear choice o

Batch-of-Thought: Cross-Instance Learning for Enhanced LLM Reasoning

AgentsDGX agent

arXiv:2601.02950v3 Announce Type: replace Abstract: Current Large Language Model reasoning systems process queries independently, discarding valuable cross-instance signals such as shared reasoning pa

BEACON: A Multimodal Dataset for Learning Behavioral Fingerprints from Gameplay Data

Model ReleasesDGX agent

arXiv:2605.10867v1 Announce Type: cross Abstract: Continuous authentication in high-stakes digital environments requires datasets with fine-grained behavioral signals under realistic cognitive and mot

Behavioral Determinants of Deployed AI Agents in Social Networks: A Multi-Factor Study of Personality, Model, and Guardrail Specification

SafetyDGX agent

arXiv:2605.08463v1 Announce Type: new Abstract: Autonomous AI agents are increasingly deployed in open social environments, yet the relationship between their configuration specifications and their em

Belief or Circuitry? Causal Evidence for In-Context Graph Learning

Local AiDGX agent

arXiv:2605.08405v1 Announce Type: new Abstract: How do LLMs learn in-context? Is it by pattern-matching recent tokens, or by inferring latent structure? We probe this question using a toy graph random

BenchCAD: A Comprehensive, Industry-Standard Benchmark for Programmatic CAD

Model ReleasesDGX agent

arXiv:2605.10865v1 Announce Type: new Abstract: Industrial Computer-Aided Design (CAD) code generation requires models to produce executable parametric programs from visual or textual inputs. Beyond r

Benchmarking Compositional Generalisation for Machine Learning Interatomic Potentials

Model ReleasesDGX agent

arXiv:2605.08988v1 Announce Type: cross Abstract: Machine Learning Interatomic Potentials play a fundamental role in computational chemistry and materials science, enabling applications from molecular

Benchmarking ResNet Backbones in RT-DETR: Impact of Depth and Regularization under environmental conditions

ResearchDGX agent

arXiv:2605.08136v1 Announce Type: cross Abstract: Visual perception plays a central role in competitive robotics, where environmental variations can directly affect real-time detection performance. Th

Benchmarking Safety Risks of Knowledge-Intensive Reasoning under Malicious Knowledge Editing

Model ReleasesDGX agent

arXiv:2605.10146v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on knowledge editing to support knowledge-intensive reasoning, but this flexibility also introduces criti

Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis

ResearchDGX agent

arXiv:2407.12173v2 Announce Type: replace-cross Abstract: Generative diffusion models have emerged as a powerful tool for high-quality image synthesis, yet their iterative nature demands significant c

Beyond Accuracy: Evaluating Strategy Diversity in LLM Mathematical Reasoning

Model ReleasesDGX agent

arXiv:2605.09292v1 Announce Type: new Abstract: Large language models now achieve high final-answer accuracy on mathematical reasoning benchmarks, but accuracy alone does not capture reasoning flexibi

Beyond Autonomy: A Dynamic Tiered AgentRunner Framework for Governable and Resilient Enterprise AI Execution

SafetyDGX agent

arXiv:2605.10223v1 Announce Type: new Abstract: Current large language model agent frameworks prioritize autonomy but lack the governability mechanisms required for enterprise deployment. High-risk wr

Beyond Continuity: Challenges of Context Switching in Multi-Turn Dialogue with LLMs

SafetyDGX agent

arXiv:2605.09268v1 Announce Type: cross Abstract: Users interacting with Large Language Models (LLMs) in a multi-turn conversation routinely refine their requests or pivot to new topics. LLMs, however

Beyond ESG Scores: Learning Dynamic Constraints for Sequential Portfolio Optimization

SafetyDGX agent

arXiv:2605.09310v1 Announce Type: new Abstract: ESG-aware portfolio optimization is increasingly important for sustainable capital allocation, yet most learning-based methods still operationalize ESG

Beyond Penalization: Diffusion-based Out-of-Distribution Detection and Selective Regularization in Offline Reinforcement Learning

SafetyDGX agent

arXiv:2605.08202v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) faces a critical challenge of overestimating the value of out-of-distribution (OOD) actions. Existing methods miti

Beyond Self-Play: Hierarchical Reasoning for Continuous Motion in Closed-Loop Traffic Simulation

SafetyDGX agent

arXiv:2605.09153v1 Announce Type: cross Abstract: Closed-loop traffic simulation requires agents that are both scalable and behaviorally realistic. Recent self-play reinforcement learning approaches d

Beyond the False Trade-off: Adaptive EWC for Stealthy and Generalizable T2I Backdoors

Model ReleasesDGX agent

arXiv:2605.08280v1 Announce Type: cross Abstract: Preserving model fidelity is essential for stealthy text-to-image (T2I) backdoor attacks. Existing methods such as Learning without Forgetting (LwF) r

Beyond the Last Layer: Multi-Layer Representation Fusion for Visual Tokenizatio

ResearchDGX agent

arXiv:2605.10780v1 Announce Type: cross Abstract: Representation autoencoders that reuse frozen pretrained vision encoders as visual tokenizers have achieved strong reconstruction and generation quali

Beyond the Singular: Revealing the Value of Multiple Generations in Benchmark Evaluation

Model ReleasesDGX agent

arXiv:2502.08943v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated significant utility in real-world applications, exhibiting impressive capabilities in natural l

Bias by Necessity: Impossibility Theorems for Sequential Processing with Convergent AI and Human Validation

SafetyDGX agent

arXiv:2605.08716v1 Announce Type: new Abstract: Are certain cognitive biases mathematically inevitable consequences of sequential information processing? We prove that primacy effects, anchoring, and

Big AI is accelerating the metacrisis: What can we do?

SafetyDGX agent

arXiv:2512.24863v2 Announce Type: replace-cross Abstract: The world is in the grip of ecological, meaning, and language crises that are converging into a metacrisis. Big AI is accelerating them all. L

Biological Plausibility and Representational Alignment of Feedback Alignment in Convolutional Networks

SafetyDGX agent

arXiv:2605.08564v1 Announce Type: new Abstract: The feedback alignment (FA) algorithm offers a biologically plausible alternative to backpropagation (BP) for training neural networks yet notably fails

Biosignal Fingerprinting: A Cross-Modal PPG-ECG Foundation Model

ApplicationsDGX agent

arXiv:2605.09579v1 Announce Type: cross Abstract: Cardiovascular disease remains the leading cause of global mortality, yet scalable cardiac monitoring is hindered by the gap between diagnostic-rich E

BONSAI: Bayesian Optimization with Natural Simplicity and Interpretability

SafetyDGX agent

arXiv:2602.07144v2 Announce Type: replace-cross Abstract: Bayesian optimization (BO) is a popular technique for sample-efficient optimization of black-box functions. In many applications, the paramete

BoostAPR: Boosting Automated Program Repair via Execution-Grounded Reinforcement Learning with Dual Reward Models

ResearchDGX agent

arXiv:2605.09134v1 Announce Type: new Abstract: Reinforcement learning for program repair is hindered by sparse execution feedback and coarse sequence-level rewards that obscure which edits actually f

Break the Brake, Not the Wheel: Untargeted Jailbreak via Entropy Maximization

SafetyDGX agent

arXiv:2605.10764v1 Announce Type: cross Abstract: Recent studies show that gradient-based universal image jailbreaks on vision-language models (VLMs) exhibit little or no cross-model transferability,

Breaking Contextual Inertia: Reinforcement Learning with Single-Turn Anchors for Stable Multi-Turn Interaction

ResearchDGX agent

arXiv:2603.04783v2 Announce Type: replace Abstract: While LLMs demonstrate strong reasoning capabilities when provided with full information in a single turn, they exhibit substantial vulnerability in

Breaking the Grid: Distance-Guided Reinforcement Learning in Large Discrete Action Spaces

SafetyDGX agent

arXiv:2602.08616v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) is increasingly applied to large-scale decision-making problems like logistics, scheduling, and recommender system

Bridging Modalities, Spanning Time: Structured Memory for Ultra-Long Agentic Video Reasoning

Model ReleasesDGX agent

arXiv:2605.08271v1 Announce Type: cross Abstract: Understanding ultra-long videos such as egocentric recordings, live streams, or surveillance footage spanning days to weeks, remains a challenge. For

Bridging Sequence and Graph Structure for Epigenetic Age Prediction

ResearchDGX agent

arXiv:2605.10541v1 Announce Type: new Abstract: Epigenetic clocks based on DNA methylation have emerged as powerful tools for estimating biological age, with broad applications in aging research, age-

Bridging the Cognitive Gap: A Unified Memory Paradigm for 6G Agentic AI-RAN

AgentsDGX agent

arXiv:2605.10036v1 Announce Type: cross Abstract: As 6G evolves, the radio access network must transcend traditional automation to embrace agentic AI capable of perception, reasoning, and evolution. A

BubbleSpec: Turning Long-Tail Bubbles into Speculative Rollout Drafts for Synchronous Reinforcement Learning

ResearchDGX agent

arXiv:2605.08862v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has become a cornerstone for improving the performance of Large Language Models (LLMs). However, its rollout phase constit

Budget-Efficient Automatic Algorithm Design via Code Graph

SafetyDGX agent

arXiv:2605.10598v1 Announce Type: new Abstract: Large language models (LLMs) have emerged as powerful tools for automatic algorithm design (AAD). However, existing pipelines remain inefficient. They o

Built Environment Reasoning from Remote Sensing Imagery Using Large Vision--Language Models

Model ReleasesDGX agent

arXiv:2605.08404v1 Announce Type: cross Abstract: This work investigates the use of large language models (LLMs) for tasks in smart cities. The core idea is to leverage remote sensing imagery to chara

bViT: Investigating Single-Block Recurrence in Vision Transformers for Image Recognition

Model ReleasesDGX agent

arXiv:2605.10661v1 Announce Type: cross Abstract: Vision Transformers (ViTs) are built by stacking independently parameterized blocks, but it remains unclear how much of this depth requires layer spec

C2L-Net: A Data-Driven Model for State-of-Charge Estimation of Lithium-Ion Batteries During Discharge

Local AiDGX agent

arXiv:2605.08653v1 Announce Type: new Abstract: Accurate state-of-charge (SOC) estimation is critical for the safe and efficient operation of lithium-ion batteries in battery management systems (BMS).

CachePrune: Teaching LLMs What Not to Follow via KV-Cache Editing

ResearchDGX agent

arXiv:2504.21228v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are susceptible to indirect prompt injection attacks, where the model inadvertently responds to instructions inje

CADBench: A Multimodal Benchmark for AI-Assisted CAD Program Generation

Model ReleasesDGX agent

arXiv:2605.10873v1 Announce Type: cross Abstract: Recovering editable CAD programs from images or 3D observations is central to AI-assisted design, but progress is difficult to measure because existin

CalBench: Evaluating Coordination-Privacy Trade-offs in Multi-Agent LLMs

SafetyDGX agent

arXiv:2605.09823v1 Announce Type: cross Abstract: We introduce CalBench, a controlled evaluation environment for studying multi-agent coordination through calendar scheduling. In CalBench, N agents ea

CAMAL: Improving Attention Alignment and Faithfulness with Segmentation Masks

SafetyDGX agent

arXiv:2605.08325v1 Announce Type: cross Abstract: Many vision datasets now provide segmentation masks in addition to annotated images to support a wide range of tasks. In this work, we propose Class A

Can Agent Benchmarks Support Their Scores? Evidence-Supported Bounds for Interactive-Agent Evaluation

Model ReleasesDGX agent

arXiv:2605.10448v1 Announce Type: new Abstract: Interactive agent benchmarks map an agent run to a binary outcome through outcome checks. When these checks rely on surface level signals or fail to cap

Can Language Models Analyze Data? Evaluating Large Language Models for Question Answering over Datasets

ResearchDGX agent

arXiv:2605.10419v1 Announce Type: cross Abstract: This paper investigates the effectiveness of large language models (LLMs) in answering questions over datasets. We examine their performance in two sc

Can LLMs Estimate Student Struggles? Human-AI Difficulty Alignment with Proficiency Simulation for Item Difficulty Prediction

SafetyDGX agent

arXiv:2512.18880v2 Announce Type: replace-cross Abstract: Accurate estimation of item (question or task) difficulty is critical for educational assessment but suffers from the cold start problem. Whil

Can LLMs Predict Polymer Physics Just by Reading Synthesis and Processing Prose?

Model ReleasesDGX agent

arXiv:2605.08255v1 Announce Type: cross Abstract: Can large language models predict physical and mechanical polymer properties simply by reading unstructured scientific prose? Polymer performance is r

Can RL Teach Long-Horizon Reasoning to LLMs? Expressiveness Is Key

ResearchDGX agent

arXiv:2605.06638v2 Announce Type: replace Abstract: Reinforcement learning (RL) has been applied to improve large language model (LLM) reasoning, yet the systematic study of how training scales with t

Can We Formally Verify Neural PDE Surrogates? SMT Compilation of Small Fourier Neural Operators

ApplicationsDGX agent

arXiv:2605.08938v1 Announce Type: new Abstract: Fourier Neural Operators (FNOs) can greatly accelerate PDE simulation, but they are often used without formal guarantees that they preserve basic physic

Can You Keep a Secret? Involuntary Information Leakage in Language Model Writing

ResearchDGX agent

arXiv:2605.10794v1 Announce Type: cross Abstract: Language models are deployed in settings that require compartmentalization: system prompts should not be disclosed, chain-of-thought reasoning is hidd

Capacity-Aware Inference: Mitigating the Straggler Effect in Mixture of Experts

Model ReleasesDGX agent

arXiv:2503.05066v5 Announce Type: replace-cross Abstract: The Mixture of Experts (MoE) is an effective architecture for scaling large language models by leveraging sparse expert activation to balance

CARL: Criticality-Aware Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2512.04949v3 Announce Type: replace-cross Abstract: Agents capable of accomplishing complex tasks through multiple interactions with the environment have emerged as a popular research direction.

CATO: Charted Attention for Neural PDE Operators

ResearchDGX agent

arXiv:2605.09016v1 Announce Type: new Abstract: Neural operators have emerged as powerful data-driven solvers for PDEs, offering substantial acceleration over classical numerical methods. However, exi

Causal Dimensionality of Transformer Representations: Measurement, Scaling, and Layer Structure

Model ReleasesDGX agent

arXiv:2605.08740v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) decompose transformer residual streams into interpretable feature dictionaries, yet the relationship between SAE width and

Causal Discovery for Irregularly Time Series with Consistency Guarantees

ApplicationsDGX agent

arXiv:2507.03310v2 Announce Type: replace-cross Abstract: This paper studies causal discovery in irregularly sampled time series-a key challenge in risk-sensitive domains like finance, healthcare, and

Causal Parametric Drift Simulation: A Digital Twin Framework for Classifier Robustness Evaluation

ResearchDGX agent

arXiv:2605.09663v1 Announce Type: cross Abstract: Machine learning classifiers in dynamic environments face concept drift -- changes in the data-generating process that degrade performance. Convention

Causal Stories from Sensor Traces: Auditing Epistemic Overreach in LLM-Generated Personal Sensing Explanations

Model ReleasesDGX agent

arXiv:2605.08590v1 Announce Type: cross Abstract: LLMs are increasingly used to explain personal sensing data, translating traces of activity and mood into natural-language accounts of why an anomalou

CauSim: Scaling Causal Reasoning with Increasingly Complex Causal Simulators

TutorialsDGX agent

arXiv:2605.09079v1 Announce Type: new Abstract: Despite surpassing human performance across mathematics, coding, and other knowledge-intensive tasks, large language models (LLMs) continue to struggle

Cavity-Enhanced Collective Quantum Processing with Polarization-Encoded Qubits

Model ReleasesDGX agent

arXiv:2605.10473v1 Announce Type: cross Abstract: We introduce a cavity-enhanced optical architecture for collective quantum processing in which logical qubits are encoded in the polarization subspace

← Previous
1…263264265266267…358
Next →