AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Agents

Agentic Literacy Debt: A Structural Problem the AI Literacy Field Has Not Yet Named

DGX agent

arXiv:2605.27396v1 Announce Type: cross Abstract: Autonomous AI agents now plan, decide, and act on behalf of users across healthcare, financial services, and workplace contexts, often without step-by

agentsarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Agyn: An Open-Source Platform for AI Agents with Scalable On-Demand Execution, Agent Definition as a Code, and Zero-Trust Access

DGX agent

arXiv:2605.27575v1 Announce Type: new Abstract: As organizations move toward production deployments of AI agents, which execute non-deterministic workflows, maintain stateful sessions, and often opera

agentsarxiv-cs-ai
28 May 2026
Applications

AI in the Workplace: The Impact of AI on Perceived Job Decency and Meaningfulness

DGX agent

arXiv:2605.28680v1 Announce Type: cross Abstract: The proliferation of Artificial Intelligence (AI) in workplaces is transforming how we work. While existing research on human-AI collaboration at work

applicationsarxiv-cs-ai
28 May 2026
Safety

AI, Take the Wheel: What Drives Delegation and Trust in Human-Computer Cooperative Question Answering?

DGX agent

arXiv:2605.28255v1 Announce Type: new Abstract: AI systems are fallible, and humans can make mistakes in deciding whether to trust AI over their own judgment. Thus, improving human-AI collaboration re

safetyarxiv-cs-ai
28 May 2026
Agents

AIBuildAI-2: A Knowledge-Enhanced Agent for Automatically Building AI Models

DGX agent

arXiv:2605.27873v1 Announce Type: new Abstract: AI models underpin data-centric applications from image and text processing to scientific discovery in biology, physics, and chemistry. Yet developing t

agentsarxiv-cs-ai
28 May 2026
Model Releases

Aligning Language Model Benchmarks with Pairwise Preferences

DGX agent

arXiv:2602.02898v2 Announce Type: replace Abstract: Language model benchmarks are pervasive and computationally-efficient proxies for real-world performance. However, many recent works find that bench

model-releasesarxiv-cs-ai
28 May 2026
Tutorials

Aligning LLMs with Human Uncertainty: A Beta-Bernoulli Calibrator for LLM Forecasting

DGX agent

arXiv:2605.27668v1 Announce Type: cross Abstract: Probabilistic forecasting estimates the likelihood of uncertain future events. To improve LLM forecasting, existing methods typically learn from binar

tutorialsarxiv-cs-ai
28 May 2026
Model Releases

AlphaForgeBench: Benchmarking End-to-End Trading Strategy Design with Large Language Models

DGX agent

arXiv:2602.18481v2 Announce Type: replace-cross Abstract: The rapid advancement of Large Language Models (LLMs) has led to a surge of financial benchmarks, evolving from static knowledge evaluation to

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

AlphaTransit: Learning to Design City-scale Transit Routes

DGX agent

arXiv:2605.28730v1 Announce Type: new Abstract: Designing a transit network requires many sequential route extension decisions, but their quality is often visible only after the full network is assemb

model-releasesarxiv-cs-ai
28 May 2026
Research

An Empirical Audit of k-NAF Budget Accounting for Anchored Decoding

DGX agent

arXiv:2605.28001v1 Announce Type: new Abstract: We empirically audit the k-NAF budget-accounting mechanism in Anchored Decoding using (i) a fixed, class-stratified workload (approximately 8,500 random

researcharxiv-cs-ai
28 May 2026
Model Releases

An Enhanced Large Neighborhood Search Approach for the Capacitated Facility Location Problem with Incompatible Customers

DGX agent

arXiv:2605.28337v1 Announce Type: new Abstract: A new variant of the classic capacitated facility location problem, which considers incompatibilities between customers, has recently been introduced in

model-releasesarxiv-cs-ai
28 May 2026
Agents

An LLM-Based Assistance System for Intuitive and Flexible Capability-Based Planning

DGX agent

arXiv:2605.28666v1 Announce Type: new Abstract: In modern industry, dynamic environments and the complexity of modular and reconfigurable resources require automated planning of process sequences. Cap

agentsarxiv-cs-ai
28 May 2026
Research

Anomaly as Non-Conformity via Training-Free Graph Laplacian Energy Minimization

DGX agent

arXiv:2605.28428v1 Announce Type: cross Abstract: Detecting subtle visual anomalies in images remains challenging, particularly when only normal samples are available a priori. Such unsupervised anoma

researcharxiv-cs-ai
28 May 2026
Model Releases

Apple Intelligence Foundation Language Models

DGX agent

arXiv:2407.21075v2 Announce Type: replace Abstract: We present foundation language models developed to power Apple Intelligence features, including a ~3 billion parameter model designed to run efficie

model-releasesarxiv-cs-ai
28 May 2026
Applications

Architecture-driven Shift: towards a lightweight selector for capturing the trends of logit shift

DGX agent

arXiv:2605.27469v1 Announce Type: cross Abstract: Continual Learning (CL) is a practical paradigm to utilize power of deep pre-trained neural networks, but which pre-trained model has a better ability

applicationsarxiv-cs-ai
28 May 2026
Research

Asking Is Not Enough: Protocol Sensitivity in LLM Confidence Calibration

DGX agent

arXiv:2605.27752v1 Announce Type: new Abstract: LLM confidence calibration is often evaluated by comparing two signals: token-probability scores and verbalized confidence. These signals are sometimes

researcharxiv-cs-ai
28 May 2026
Model Releases

AssertLLM2: A Comprehensive LLM Benchmark for Assertion Generation from Design Specifications

DGX agent

arXiv:2605.27472v1 Announce Type: cross Abstract: Assertion-based verification (ABV) is a cornerstone of modern hardware design, yet manually translating design intent into formal SystemVerilog Assert

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

ASTRA: Communication-Efficient Acceleration for Multi-Device Transformer Inference

DGX agent

arXiv:2505.19342v2 Announce Type: replace-cross Abstract: Multi-device inference can reduce Transformer latency by parallelizing computation. However, existing methods require high inter-device bandwi

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios

DGX agent

arXiv:2605.27995v1 Announce Type: new Abstract: Large language model (LLM)-based agents have shown strong capabilities in using external tools to solve complex tasks. However, existing evaluations oft

model-releasesarxiv-cs-ai
28 May 2026
Research

Atomic Skills are the Prerequisite: When Reinforcement Learning Synthesizes Compositional Reasoning, and When It Only Amplifies

DGX agent

arXiv:2512.01970v3 Announce Type: replace Abstract: Does Reinforcement Learning (RL) merely amplify existing skills, or synthesize novel skills? We investigate this question through the lens of Comple

researcharxiv-cs-ai
28 May 2026
Safety

Auditable Decision Models with Learned Abstention and Real-Time Steering

DGX agent

arXiv:2605.27768v1 Announce Type: new Abstract: Production AI systems often operate with incomplete, conflicting, or insufficient evidence. Forced classifiers collapse such cases into action labels, w

safetyarxiv-cs-ai
28 May 2026
Agents

AutoScientists: Self-Organizing Agent Teams for Long-Running Scientific Experimentation

DGX agent

arXiv:2605.28655v1 Announce Type: new Abstract: Scientific research proceeds through iterative cycles of hypothesis generation, experiment design, execution, and revision. AI agents can automate parts

agentsarxiv-cs-ai
28 May 2026
Local Ai

Backdoor Attacks on Fault Detection and Localization in Cyber-Physical Systems

DGX agent

arXiv:2605.27674v1 Announce Type: cross Abstract: Cyber-Physical Systems (CPS) integrate sensing, communication, computation, and control to support critical infrastructure, including smart grids, ind

local-aiarxiv-cs-ai
28 May 2026
Research

Balancing Fidelity and Diversity in Diffusion Models via Symmetric Attention Decomposition: Hopfield Perspective

DGX agent

arXiv:2605.27476v1 Announce Type: cross Abstract: We characterize the pre-softmax attention matrix mathbf{QK^op} in transformers as an associative memory matrix encoding pairwise associations between

researcharxiv-cs-ai
28 May 2026
Model Releases

Bandwidth-Efficient and Privacy-Preserving Edge-Cloud Many-to-Many Speech Translation

DGX agent

arXiv:2605.28642v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have demonstrated significant potential for speech-to-text translation (S2TT). However, existing deployment par

model-releasesarxiv-cs-ai
28 May 2026
Safety

Bayesian Gated Non-Negative Contrastive Learning

DGX agent

arXiv:2605.28441v1 Announce Type: cross Abstract: While Contrastive Learning (CL) has revolutionized self-supervised representation learning, its latent representations remain highly entangled and opa

safetyarxiv-cs-ai
28 May 2026
Safety

Behavioural Analysis of Alignment Faking

DGX agent

arXiv:2605.27681v1 Announce Type: new Abstract: Alignment faking (AF) refers to a model strategically complying with a training objective to avoid behavioural modification while preserving its deploym

safetyarxiv-cs-ai
28 May 2026
Model Releases

Benchmarking AI for low-resource contexts: Thinking beyond leaderboards

DGX agent

arXiv:2605.28508v1 Announce Type: new Abstract: Existing AI evaluation practices often fail to capture how systems actually perform in low-resource environments, where operational constraints shape us

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Benchmarking Fairness in Spiking Neural Networks: Data Bias, Spurious Features, and Hardware Effects

DGX agent

arXiv:2605.27407v1 Announce Type: cross Abstract: Evaluating fairness in Spiking Neural Networks (SNNs) demands rigorous benchmarks that reflect real-world complexities, yet existing assessments remai

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Benchmarks are Not Enough: RAMP for Runtime Assessing of Agentic Models in Production Systems

DGX agent

arXiv:2605.27492v1 Announce Type: cross Abstract: LLM agents are rapidly evolving from coding assistants into autonomous software engineering systems. However, existing evaluation methodologies remain

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

BenGER: Benchmarking LLM Systems on Subsumption-Based Legal Reasoning in German Law

DGX agent

arXiv:2605.28183v1 Announce Type: cross Abstract: We introduce the BenGER (Benchmark for German Law) dataset for evaluating LLM systems on subsumption-based legal reasoning in German law. The BenGER d

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Better Accuracies, Worse Reasoning: A Step-Level Audit of Medical Chain-of-Thought Distillation

DGX agent

arXiv:2605.28301v1 Announce Type: new Abstract: Chain-of-thought (CoT) distillation trains a smaller model to imitate a teacher's reasoning trace, but it is typically evaluated by final-answer metrics

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Beyond Binary Moral Judgment: Modeling Ethical Pluralism in AI

DGX agent

arXiv:2605.28707v1 Announce Type: new Abstract: Critical decision-making in socially consequential spaces is increasingly involving AI systems at varying capacities. Yet, despite the ubiquity of auton

model-releasesarxiv-cs-ai
28 May 2026
Safety

Beyond Binary: Sim-to-Real Dexterous Manipulation with Physics-Grounded Contact Representation

DGX agent

arXiv:2605.28812v1 Announce Type: cross Abstract: A primary bottleneck in contact-rich manipulation is the difficulty of collecting real-world data. Sim-to-real reinforcement learning offers a scalabl

safetyarxiv-cs-ai
28 May 2026
Model Releases

Beyond External Monitors: Enhancing Transparency of Large Language Models for Easier Monitoring

DGX agent

arXiv:2502.05242v3 Announce Type: replace-cross Abstract: Large language models (LLMs) are becoming increasingly capable, but the mechanisms of their thinking and decision-making processes remain uncl

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Beyond Model Ranking: Predictability-Aligned Evaluation for Time Series Forecasting

DGX agent

arXiv:2509.23074v3 Announce Type: replace-cross Abstract: In the era of increasingly complex AI models for time series forecasting, progress is often measured by marginal improvements on benchmark lea

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

Beyond Motion Primitives: Behavioral Activity Recognition from Head-Mounted IMU

DGX agent

arXiv:2605.27464v1 Announce Type: cross Abstract: AR smart glasses need continuous behavioral context to offer proactive assistance, yet their most practical always-on sensor, the head-mounted Inertia

model-releasesarxiv-cs-ai
28 May 2026
Safety

BiasEdit: A Training-Free Bias-Detect-and-Edit Framework for Learning Fair Visual Classifiers

DGX agent

arXiv:2605.28450v1 Announce Type: cross Abstract: Visual data from the Web power image classifiers, which often underpin many web services, such as recommendation and content moderation. However, the

safetyarxiv-cs-ai
28 May 2026
Model Releases

BioELX: Cross-lingual Biomedical Entity Linking via Alias-based Retrieval and LLM Ranking

DGX agent

arXiv:2605.27380v1 Announce Type: cross Abstract: Cross-lingual biomedical entity linking (BEL) maps mentions in any language to unique identifiers in a biomedical knowledge base (KB), supporting clin

model-releasesarxiv-cs-ai
28 May 2026
Research

BIRDNet: Mining and Encoding Boolean Implication Knowledge Graphs as Interpretable Deep Neural Networks

DGX agent

arXiv:2605.28739v1 Announce Type: cross Abstract: Tabular data in knowledge-rich domains often carries a latent prior in the form of Boolean implication relationships (BIRs) between pairs of features.

researcharxiv-cs-ai
28 May 2026
Research

BIRDS: Characterizing and Understanding Biodiversity Impact of Large Language Model Serving

DGX agent

arXiv:2605.27480v1 Announce Type: cross Abstract: Large language model (LLM) serving creates environmental impacts beyond carbon and water, including ecosystem damage through biodiversity-related path

researcharxiv-cs-ai
28 May 2026
Model Releases

BlazeEdit: Generalist Image Editing on Mobile Devices with Image-to-Image Diffusion Models

DGX agent

arXiv:2605.28067v1 Announce Type: new Abstract: The remarkable generation quality of modern diffusion models often comes at the cost of massive parameter counts, which necessitate server-side inferenc

model-releasesarxiv-cs-ai
28 May 2026
Safety

Blind PRNG Hijacking: An Undetectable Integrity-Preserving Attack Against LLM Watermarking

DGX agent

arXiv:2605.28632v1 Announce Type: cross Abstract: Cryptographic watermarking is a leading defense for attributing text generated by large language models (LLMs). Existing schemes, including KGW, Unigr

safetyarxiv-cs-ai
28 May 2026
Safety

Bridging the Detection-to-Abstention Gap in Reasoning Models under Insufficient Information

DGX agent

arXiv:2605.28070v1 Announce Type: new Abstract: We highlight a failure mode of large reasoning models on questions with insufficient information: models may recognize that a problem is under-specified

safetyarxiv-cs-ai
28 May 2026
Model Releases

Bridging the Stability-Expressivity Gap: Synthetic Data Scaling and Preference Alignment for Low-Resource Spoken Language Models

DGX agent

arXiv:2605.27383v1 Announce Type: cross Abstract: Spoken Language Models (SLMs) have emerged as a promising paradigm for speech synthesis by bypassing explicit grapheme-to-phoneme pipelines. However,

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

BuddyBench: A Privacy-Constrained Multi-Task Benchmark for Pediatric Social-Communication Personalization

DGX agent

arXiv:2605.28089v1 Announce Type: new Abstract: BuddyBench introduces a privacy-constrained multi-task benchmark for pediatric social-communication personalization. Unlike existing neurodevelopmental

model-releasesarxiv-cs-ai
28 May 2026
Tutorials

C-MIG: Multi-view Information Gain-based Retrieval-Augmented Generation for Clinical Diagnosis Reasoning

DGX agent

arXiv:2605.27860v1 Announce Type: new Abstract: Retrieval-augmented generation combined with reinforcement learning has shown promise for grounding large language models in trustworthy medical evidenc

tutorialsarxiv-cs-ai
28 May 2026
Agents

Calibrating Conservatism for Scalable Oversight

DGX agent

arXiv:2605.28807v1 Announce Type: new Abstract: Agentic AI systems capable of autonomous planning and extended environmental interaction pose a fundamental control problem: how can humans maintain mea

agentsarxiv-cs-ai
28 May 2026
← Previous
1…245246247248249…448
Next →