AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlog
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Research

Detecting Data Contamination in Large Language Models

DGX agent

arXiv:2604.19561v1 Announce Type: new Abstract: Large Language Models (LLMs) utilize large amounts of data for their training, some of which may come from copyrighted sources. Membership Inference Att

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Detecting Hallucinations in SpeechLLMs at Inference Time Using Attention Maps

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2604.19565v1 Announce Type: cross Abstract: Hallucinations in Speech Large Language Models (SpeechLLMs) pose significant risks, yet existing detection methods typically rely on gold-standard out

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Diamond Maps: Efficient Reward Alignment via Stochastic Flow Maps

DGX agent

arXiv:2602.05993v2 Announce Type: replace-cross Abstract: Flow and diffusion models produce high-quality samples, but adapting them to user preferences or constraints post-training remains costly and

safetyarxiv-cs-ai
22 Apr 2026
Local Ai

Distillation Traps and Guards: A Calibration Knob for LLM Distillability

DGX agent

arXiv:2604.18963v1 Announce Type: cross Abstract: Knowledge distillation (KD) transfers capabilities from large language models (LLMs) to smaller students, yet it can fail unpredictably and also under

local-aiarxiv-cs-ai
22 Apr 2026
Model Releases

Do Agents Dream of Root Shells? Partial-Credit Evaluation of LLM Agents in Capture The Flag Challenges

DGX agent

arXiv:2604.19354v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly proposed for autonomous cybersecurity tasks, but their capabilities in realistic offensive settings r

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Do LLMs Game Formalization? Evaluating Faithfulness in Logical Reasoning

DGX agent

arXiv:2604.19459v1 Announce Type: new Abstract: Formal verification guarantees proof validity but not formalization faithfulness. For natural-language logical reasoning, where models construct axiom s

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

DP-FlogTinyLLM: Differentially private federated log anomaly detection using Tiny LLMs

DGX agent

arXiv:2604.19118v1 Announce Type: cross Abstract: Modern distributed systems generate massive volumes of log data that are critical for detecting anomalies and cyber threats. However, in real world se

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

DT2IT-MRM: Debiased Preference Construction and Iterative Training for Multimodal Reward Modeling

DGX agent

arXiv:2604.19544v1 Announce Type: new Abstract: Multimodal reward models (MRMs) play a crucial role in aligning Multimodal Large Language Models (MLLMs) with human preferences. Training a good MRM req

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

DW-Bench: Benchmarking LLMs on Data Warehouse Graph Topology Reasoning

DGX agent

arXiv:2604.18964v1 Announce Type: new Abstract: This paper introduces DW-Bench, a new benchmark that evaluates large language models (LLMs) on graph-topology reasoning over data warehouse schemas, exp

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Early Pruning for Public Transport Routing

DGX agent

arXiv:2603.12592v2 Announce Type: replace-cross Abstract: Routing algorithms for public transport, particularly the widely used RAPTOR and its variants, often face performance bottlenecks during the t

researcharxiv-cs-ai
22 Apr 2026
Research

Easy Samples Are All You Need: Self-Evolving LLMs via Data-Efficient Reinforcement Learning

DGX agent

arXiv:2604.18639v1 Announce Type: cross Abstract: Previous LLMs-based RL studies typically follow either supervised learning with high annotation costs, or unsupervised paradigms using voting or entro

researcharxiv-cs-ai
22 Apr 2026
Research

EgoSelf: From Memory to Personalized Egocentric Assistant

DGX agent

arXiv:2604.19564v1 Announce Type: cross Abstract: Egocentric assistants often rely on first-person view data to capture user behavior and context for personalized services. Since different users exhib

researcharxiv-cs-ai
22 Apr 2026
Research

EHRAG: Bridging Semantic Gaps in Lightweight GraphRAG via Hybrid Hypergraph Construction and Retrieval

DGX agent

arXiv:2604.17458v2 Announce Type: replace Abstract: Graph-based Retrieval-Augmented Generation (GraphRAG) enhances LLMs by structuring corpus into graphs to facilitate multi-hop reasoning. While recen

researcharxiv-cs-ai
22 Apr 2026
Applications

Enabling Vibration-Based Gesture Recognition on Everyday Furniture via Energy-Efficient FPGA Implementation of 1D Convolutional Networks

DGX agent

arXiv:2510.23156v2 Announce Type: replace-cross Abstract: The growing demand for smart home interfaces has increased interest in non-intrusive sensing methods like vibration-based gesture recognition.

applicationsarxiv-cs-ai
22 Apr 2026
Tutorials

End-to-End Large Portfolio Optimization for Variance Minimization with Neural Networks through Covariance Cleaning

DGX agent

arXiv:2507.01918v3 Announce Type: replace-cross Abstract: We develop a rotation-invariant neural network that provides the global minimum-variance portfolio by jointly learning how to lag-transform hi

tutorialsarxiv-cs-ai
22 Apr 2026
Safety

Enhancing Construction Worker Safety in Extreme Heat: A Machine Learning Approach Utilizing Wearable Technology for Predictive Health Analytics

DGX agent

arXiv:2604.19559v1 Announce Type: new Abstract: Construction workers are highly vulnerable to heat stress, yet tools that translate real-time physiological data into actionable safety intelligence rem

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Environmental Sound Deepfake Detection Using Deep-Learning Framework

DGX agent

arXiv:2604.19652v1 Announce Type: cross Abstract: In this paper, we propose a deep-learning framework for environmental sound deepfake detection (ESDD) -- the task of identifying whether the sound sce

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Epistemic Skills: Reasoning about Knowledge and Oblivion

DGX agent

arXiv:2504.01733v4 Announce Type: replace Abstract: This paper presents a class of epistemic logics that captures the dynamics of acquiring knowledge and descending into oblivion, while incorporating

researcharxiv-cs-ai
22 Apr 2026
Research

Error-free Training for MedMNIST Datasets

DGX agent

arXiv:2604.18916v1 Announce Type: new Abstract: In this paper, we introduce a new concept called Artificial Special Intelligence by which Machine Learning models for the classification problem can be

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Evaluating Answer Leakage Robustness of LLM Tutors against Adversarial Student Attacks

DGX agent

arXiv:2604.18660v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in education, yet their default helpfulness often conflicts with pedagogical principles. Prior work

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Evaluation-driven Scaling for Scientific Discovery

DGX agent

arXiv:2604.19341v1 Announce Type: cross Abstract: Language models are increasingly used in scientific discovery to generate hypotheses, propose candidate solutions, implement systems, and iteratively

local-aiarxiv-cs-ai
22 Apr 2026
Agents

EvoMaster: A Foundational Evolving Agent Framework for Agentic Science at Scale

DGX agent

arXiv:2604.17406v2 Announce Type: replace Abstract: The convergence of large language models and agents is catalyzing a new era of scientific discovery: Agentic Science. While the scientific method is

agentsarxiv-cs-ai
22 Apr 2026
Safety

EVPO: Explained Variance Policy Optimization for Adaptive Critic Utilization in LLM Post-Training

DGX agent

arXiv:2604.19485v1 Announce Type: cross Abstract: Reinforcement learning (RL) for LLM post-training faces a fundamental design choice: whether to use a learned critic as a baseline for policy optimiza

safetyarxiv-cs-ai
22 Apr 2026
Research

Experiments or Outcomes? Probing Scientific Feasibility in Large Language Models

DGX agent

arXiv:2604.18786v1 Announce Type: cross Abstract: Scientific feasibility assessment asks whether a claim is consistent with established knowledge and whether experimental evidence could support or ref

researcharxiv-cs-ai
22 Apr 2026
Safety

ExpertGen: Scalable Sim-to-Real Expert Policy Learning from Imperfect Behavior Priors

DGX agent

arXiv:2603.15956v2 Announce Type: replace-cross Abstract: Learning generalizable and robust behavior cloning policies requires large volumes of high-quality robotics data. While human demonstrations (

safetyarxiv-cs-ai
22 Apr 2026
Agents

Explicit Trait Inference for Multi-Agent Coordination

DGX agent

arXiv:2604.19278v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) show promise on complex tasks but remain prone to coordination failures such as goal drift, error cascades, and misa

agentsarxiv-cs-ai
22 Apr 2026
Safety

Failure Modes in Multi-Hop QA: The Weakest Link Effect and the Recognition Bottleneck

DGX agent

arXiv:2601.12499v2 Announce Type: replace Abstract: Despite scaling to massive context windows, Large Language Models (LLMs) struggle with multi-hop reasoning due to inherent position bias, which caus

safetyarxiv-cs-ai
22 Apr 2026
Safety

Fairness Audits of Institutional Risk Models in Deployed ML Pipelines

DGX agent

arXiv:2604.19468v1 Announce Type: cross Abstract: Fairness audits of institutional risk models are critical for understanding how deployed machine learning pipelines allocate resources. Drawing on mul

safetyarxiv-cs-ai
22 Apr 2026
Safety

FASE : A Fairness-Aware Spatiotemporal Event Graph Framework for Predictive Policing

DGX agent

arXiv:2604.18644v1 Announce Type: cross Abstract: Predictive policing systems that allocate patrol resources based solely on predicted crime risk can unintentionally amplify racial disparities through

safetyarxiv-cs-ai
22 Apr 2026
Safety

FASTER: Value-Guided Sampling for Fast RL

DGX agent

arXiv:2604.19730v1 Announce Type: cross Abstract: Some of the most performant reinforcement learning algorithms today can be prohibitively expensive as they use test-time scaling methods such as sampl

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

FedProxy: Federated Fine-Tuning of LLMs via Proxy SLMs and Heterogeneity-Aware Fusion

DGX agent

arXiv:2604.19015v1 Announce Type: cross Abstract: Federated fine-tuning of Large Language Models (LLMs) is obstructed by a trilemma of challenges: protecting LLMs intellectual property (IP), ensuring

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Fine-Tuning Code Language Models to Detect Cross-Language Bugs

DGX agent

arXiv:2507.21954v2 Announce Type: replace-cross Abstract: Multilingual programming, which involves using multiple programming languages (PLs) in a single project, is increasingly common due to its ben

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Fine-tuning DeepSeek-OCR-2 for Molecular Structure Recognition

DGX agent

arXiv:2604.03476v2 Announce Type: replace-cross Abstract: Optical Chemical Structure Recognition (OCSR) is critical for converting 2D molecular diagrams from printed literature into machine-readable f

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Fine-Tuning Small Reasoning Models for Quantum Field Theory

DGX agent

arXiv:2604.18936v1 Announce Type: cross Abstract: Despite the growing application of Large Language Models (LLMs) to theoretical physics, there is little academic exploration into how domain-specific

model-releasesarxiv-cs-ai
22 Apr 2026
Applications

Formally Verified Patent Analysis via Dependent Type Theory: Machine-Checkable Certificates from a Hybrid AI + Lean 4 Pipeline

DGX agent

arXiv:2604.18882v1 Announce Type: new Abstract: We present a formally verified framework for patent analysis as a hybrid AI + Lean 4 pipeline. The DAG-coverage core (Algorithm 1b) is fully machine-ver

applicationsarxiv-cs-ai
22 Apr 2026
Model Releases

Four-Axis Decision Alignment for Long-Horizon Enterprise AI Agents

DGX agent

arXiv:2604.19457v1 Announce Type: new Abstract: Long-horizon enterprise agents make high-stakes decisions (loan underwriting, claims adjudication, clinical review, prior authorization) under lossy mem

model-releasesarxiv-cs-ai
22 Apr 2026
Agents

From Craft to Kernel: A Governance-First Execution Architecture and Semantic ISA for Agentic Computers

DGX agent

arXiv:2604.18652v1 Announce Type: cross Abstract: The transition of agentic AI from brittle prototypes to production systems is stalled by a pervasive crisis of craft. We suggest that the prevailing o

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

From Experience to Skill: Multi-Agent Generative Engine Optimization via Reusable Strategy Learning

DGX agent

arXiv:2604.19516v1 Announce Type: new Abstract: Generative engines (GEs) are reshaping information access by replacing ranked links with citation-grounded answers, yet current Generative Engine Optimi

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

From Natural Language to Executable Narsese: A Neuro-Symbolic Benchmark and Pipeline for Reasoning with NARS

DGX agent

arXiv:2604.18873v1 Announce Type: new Abstract: Large language models (LLMs) are highly capable at language generation, but they remain unreliable when reasoning requires explicit symbolic structure,

model-releasesarxiv-cs-ai
22 Apr 2026
Research

GAIN: Multiplicative Modulation for Domain Adaptation

DGX agent

arXiv:2604.04516v2 Announce Type: replace-cross Abstract: Adapting LLMs to new domains causes forgetting because standard methods (e.g., full fine-tuning, LoRA) inject new directions into the weight s

researcharxiv-cs-ai
22 Apr 2026
Research

GAIR: Location-Aware Self-Supervised Contrastive Pre-Training with Geo-Aligned Implicit Representations

DGX agent

arXiv:2503.16683v2 Announce Type: replace-cross Abstract: Vision Transformer (ViT) has been widely used in computer vision tasks with excellent results by providing representations for a whole image o

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Gated Memory Policy

DGX agent

arXiv:2604.18933v1 Announce Type: cross Abstract: Robotic manipulation tasks exhibit varying memory requirements, ranging from Markovian tasks that require no memory to non-Markovian tasks that depend

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Generalization at the Edge of Stability

DGX agent

arXiv:2604.19740v1 Announce Type: cross Abstract: Training modern neural networks often relies on large learning rates, operating at the edge of stability, where the optimization dynamics exhibit osci

researcharxiv-cs-ai
22 Apr 2026
Model Releases

GeoLaux: A Benchmark for Evaluating MLLMs' Geometry Performance on Long-Step Problems Requiring Auxiliary Lines

DGX agent

arXiv:2508.06226v2 Announce Type: replace Abstract: Geometry problem solving (GPS) poses significant challenges for Multimodal Large Language Models (MLLMs) in diagram comprehension, knowledge applica

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Geometric Decoupling: Diagnosing the Structural Instability of Latent

DGX agent

arXiv:2604.18804v1 Announce Type: cross Abstract: Latent Diffusion Models (LDMs) achieve high-fidelity synthesis but suffer from latent space brittleness, causing discontinuous semantic jumps during e

researcharxiv-cs-ai
22 Apr 2026
Research

GOLD-BEV: GrOund and aeriaL Data for Dense Semantic BEV Mapping of Dynamic Scenes

DGX agent

arXiv:2604.19411v1 Announce Type: cross Abstract: Understanding road scenes in a geometrically consistent, scene-centric representation is crucial for planning and mapping. We present GOLD-BEV, a fram

researcharxiv-cs-ai
22 Apr 2026
Tutorials

Gradient-Based Program Synthesis with Neurally Interpreted Languages

DGX agent

arXiv:2604.18907v1 Announce Type: cross Abstract: A central challenge in program induction has long been the trade-off between symbolic and neural approaches. Symbolic methods offer compositional gene

tutorialsarxiv-cs-ai
22 Apr 2026
Safety

GRAIL:Learning to Interact with Large Knowledge Graphs for Retrieval Augmented Reasoning

DGX agent

arXiv:2508.05498v2 Announce Type: replace Abstract: Large Language Models (LLMs) integrated with Retrieval-Augmented Generation (RAG) techniques have exhibited remarkable performance across a wide ran

safetyarxiv-cs-ai
22 Apr 2026
← Previous
1…396397398399400…443
Next →