AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
28 Apr 2026

Selective Conformal Risk Control

ResearchDGX agent

arXiv:2512.12844v2 Announce Type: replace-cross Abstract: Reliable uncertainty quantification is essential for deploying machine learning systems in high-stakes domains. Conformal prediction provides

Self-Abstraction Learning for Effective and Stable Training of Deep Neural Networks

ResearchDGX agent

arXiv:2604.24313v1 Announce Type: cross Abstract: Training large-scale deep neural networks effectively and stably is essential for applying deep learning across various fields. However, conventional

Self-Admitted Technical Debt Detection Approaches: A Decade Systematic Review

ResearchDGX agent

arXiv:2312.15020v4 Announce Type: replace-cross Abstract: Technical debt (TD) refers to the long-term costs associated with suboptimal design or code decisions in software development, often made to m


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Self Knowledge Re-expression: A Fully Local Method for Adapting LLMs to Tasks Using Intrinsic Knowledge

Local AiDGX agent

arXiv:2604.22939v1 Announce Type: cross Abstract: While the next-token prediction (NTP) paradigm enables large language models (LLMs) to express their intrinsic knowledge, its sequential nature constr

SemML 2.0: Synthesizing Controllers for LTL

SafetyDGX agent

arXiv:2604.24102v1 Announce Type: new Abstract: Synthesizing a reactive system from specifications given in linear temporal logic (LTL) is a classical problem, finding its applications in safety-criti

SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning

Model ReleasesDGX agent

arXiv:2604.23747v1 Announce Type: cross Abstract: Recent mixed-policy optimization methods for LLM reasoning that interleave or blend supervised and reinforcement learning signals report improvements

SGP-SAM: Self-Gated Prompting for Transferring 3D Segment Anything Models to Lesion Segmentation

ResearchDGX agent

arXiv:2604.22825v1 Announce Type: cross Abstract: Large segmentation foundation models such as the Segment Anything Model (SAM) have reshaped promptable segmentation in natural images, and recent effo

SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning

Model ReleasesDGX agent

arXiv:2506.05425v2 Announce Type: replace-cross Abstract: Understanding social interaction, which encompasses perceiving numerous and subtle multimodal cues, inferring unobservable mental states and r

SketchVLM: Vision language models can annotate images to explain thoughts and guide users

Model ReleasesDGX agent

arXiv:2604.22875v1 Announce Type: cross Abstract: When answering questions about images, humans naturally point, label, and draw to explain their reasoning. In contrast, modern vision-language models

Skill Retrieval Augmentation for Agentic AI

Model ReleasesDGX agent

arXiv:2604.24594v1 Announce Type: cross Abstract: As large language models (LLMs) evolve into agentic problem solvers, they increasingly rely on external, reusable skills to handle tasks beyond their

Small Language Model Helps Resolve Semantic Ambiguity of LLM Prompt

ResearchDGX agent

arXiv:2604.23263v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly utilized in various complex reasoning tasks due to their excellent instruction following capability. How

SMP: Reusable Score-Matching Motion Priors for Physics-Based Character Control

SafetyDGX agent

arXiv:2512.03028v3 Announce Type: replace-cross Abstract: Data-driven motion priors that can guide agents toward producing naturalistic behaviors play a pivotal role in creating life-like virtual char

SMSI: System Model Security Inference: Automated Threat Modeling for Cyber-Physical Systems

Model ReleasesDGX agent

arXiv:2604.23905v1 Announce Type: cross Abstract: Threat modeling for cyber-physical systems (CPS) remains a largely manual exercise. This project presents SMSI (System Model Security Inference), a hy

SoccerRef-Agents: Multi-Agent System for Automated Soccer Refereeing

Model ReleasesDGX agent

arXiv:2604.23392v1 Announce Type: new Abstract: Refereeing is vital in sports, where fair, accurate, and explainable decisions are fundamental. While intelligent assistant technologies are being widel

SolarTformer: A Transformer Based Deep Learning Approach for Short Term Solar Power Forecasting

ResearchDGX agent

arXiv:2604.24306v1 Announce Type: cross Abstract: Accurate forecasting of solar power output is essential for efficient integration of renewable energy into the grid. In this study, an attention-based

Speech Enhancement Based on Drifting Models

Model ReleasesDGX agent

arXiv:2604.24199v1 Announce Type: cross Abstract: We propose Speech Enhancement based on Drifting Models (DriftSE), a novel generative framework that formulates denoising as an equilibrium problem. Ra

Speech-FT: Merging Pre-trained And Fine-Tuned Speech Representation Models For Cross-Task Generalization

Model ReleasesDGX agent

arXiv:2502.12672v4 Announce Type: replace-cross Abstract: Fine-tuning speech representation models can enhance performance on specific tasks but often compromises their cross-task generalization abili

Sphere-Depth: A Benchmark for Depth Estimation Methods with Varying Spherical Camera Orientations

Model ReleasesDGX agent

arXiv:2604.23432v1 Announce Type: cross Abstract: Reliable depth estimation from spherical images is crucial for 360{eg} vision in robotic navigation and immersive scene understanding. However, the on

SPLIT: Separating Physical-Contact via Latent Arithmetic in Image-Based Tactile Sensors

ResearchDGX agent

arXiv:2604.24449v1 Announce Type: cross Abstract: Training machine learning models for robotic tactile sensing requires vast amounts of data, yet obtaining realistic interaction data remains a challen

SSR-Zero: Simple Self-Rewarding Reinforcement Learning for Machine Translation

Model ReleasesDGX agent

arXiv:2505.16637v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated remarkable capabilities in machine translation (MT). However, most advanced MT-specifi

Statistically-Guided Meta-Learning for Cross-Deployment Activity Recognition in Distributed Fiber-Optic Sensing

ApplicationsDGX agent

arXiv:2511.17902v3 Announce Type: replace-cross Abstract: Distributed Fiber Optic Sensing (DFOS) is promising for long-range perimeter security, yet practical deployment faces three key obstacles: sev

STELLAR-E: a Synthetic, Tailored, End-to-end LLM Application Rigorous Evaluator

ResearchDGX agent

arXiv:2604.24544v1 Announce Type: new Abstract: The increasing reliance on Large Language Models (LLMs) across diverse sectors highlights the need for robust domain-specific and language-specific eval

Stochastic KV Routing: Enabling Adaptive Depth-Wise Cache Sharing

ResearchDGX agent

arXiv:2604.22782v1 Announce Type: cross Abstract: Serving transformer language models with high throughput requires caching Key-Values (KVs) to avoid redundant computation during autoregressive genera

StoryTR: Narrative-Centric Video Temporal Retrieval with Theory of Mind Reasoning

Model ReleasesDGX agent

arXiv:2604.23198v1 Announce Type: new Abstract: Current video moment retrieval excels at action-centric tasks but struggles with narrative content. Models can see extit{what is happening} but fail to

Strategic Bidding in 6G Spectrum Auctions with Large Language Models

Model ReleasesDGX agent

arXiv:2604.24156v1 Announce Type: cross Abstract: Efficient and fair spectrum allocation is a central challenge in 6G networks, where massive connectivity and heterogeneous services continuously compe

StratRAG: A Multi-Hop Retrieval Evaluation Dataset for Retrieval-Augmented Generation Systems

Model ReleasesDGX agent

arXiv:2604.22757v1 Announce Type: cross Abstract: We introduce StratRAG, an open-source retrieval evaluation dataset for benchmarking Retrieval-Augmented Generation (RAG) systems on multi-hop reasonin

Structural Enforcement of Goal Integrity in AI Agents via Separation-of-Powers Architecture

SafetyDGX agent

arXiv:2604.23646v1 Announce Type: new Abstract: Recent evidence suggests that frontier AI systems can exhibit agentic misalignment, generating and executing harmful actions derived from internally con

Structure Guided Retrieval-Augmented Generation for Factual Queries

TutorialsDGX agent

arXiv:2604.22843v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) has been proposed to mitigate hallucinations in large language models (LLMs), where generated outputs may be fact

Survey in Characterizing Semantic Change

TutorialsDGX agent

arXiv:2402.19088v5 Announce Type: replace-cross Abstract: Live languages continuously evolve to integrate the cultural change of human societies. This evolution manifests through neologisms (new words

SwarmDrive: Semantic V2V Coordination for Latency-Constrained Cooperative Autonomous Driving

AgentsDGX agent

arXiv:2604.22852v1 Announce Type: cross Abstract: Cloud-hosted LLM inference for autonomous driving adds round-trip delay and depends on stable connectivity, while purely local edge models struggle un

SycoPhantasy: Quantifying Sycophancy and Hallucination in Small Open Weight VLMs for Vision-Language Scoring of Fantasy Characters

Model ReleasesDGX agent

arXiv:2604.24346v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly deployed as evaluators in tasks requiring nuanced image understanding, yet their reliability in scoring

Symmetric Equilibrium Propagation for Thermodynamic Diffusion Training

Model ReleasesDGX agent

arXiv:2604.23806v1 Announce Type: cross Abstract: The reverse process in score-based diffusion models is formally equivalent to overdamped Langevin dynamics in a time-dependent energy landscape. In ou

SynthPert: Enhancing LLM Biological Reasoning via Synthetic Reasoning Traces for Cellular Perturbation Prediction

Model ReleasesDGX agent

arXiv:2509.25346v2 Announce Type: replace Abstract: Predicting cellular responses to genetic perturbations represents a fundamental challenge in systems biology, critical for advancing therapeutic dis

TACO: Efficient Communication Compression of Intermediate Tensors for Scalable Tensor-Parallel LLM Training

Model ReleasesDGX agent

arXiv:2604.24088v1 Announce Type: cross Abstract: Handling communication overhead in large-scale tensor-parallel training remains a critical challenge due to the dense, near-zero distributions of inte

Talking Slide Avatars: Open-Source Multimodal Communication Approach for Teaching

ApplicationsDGX agent

arXiv:2604.23703v1 Announce Type: cross Abstract: Slide-based teaching is widely used in higher education, yet in online, hybrid, and asynchronous contexts, slides often lose the instructor presence,

Tandem: Riding Together with Large and Small Language Models for Efficient Reasoning

TutorialsDGX agent

arXiv:2604.23623v1 Announce Type: new Abstract: Recent advancements in large language models (LLMs) have catalyzed the rise of reasoning-intensive inference paradigms, where models perform explicit st

Task-guided Spatiotemporal Network with Diffusion Augmentation for EEG-based Dementia Diagnosis and MMSE Prediction

ResearchDGX agent

arXiv:2604.23964v1 Announce Type: cross Abstract: Patients with dementia typically exhibit cognitive impairment, which is routinely assessed using the Mini-Mental State Examination (MMSE). Concurrentl

TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents

SafetyDGX agent

arXiv:2604.24005v1 Announce Type: cross Abstract: On-policy distillation (OPD) has shown strong potential for transferring reasoning ability from frontier or domain-specific models to smaller students

TeachMaster: Generative Teaching via Code

AgentsDGX agent

arXiv:2601.04204v2 Announce Type: replace-cross Abstract: The scalability of high-quality online education is hindered by the high costs and slow cycles of manual content creation. Despite advancement

Test of Time: Rethinking Temporal Signal of Benchmark Contamination

Model ReleasesDGX agent

arXiv:2509.00072v3 Announce Type: replace Abstract: Post-cutoff performance decay has been widely interpreted as a temporal signal for benchmark contamination. We critically examine this belief and de

Text Tells the Cost: Predicting and Analyzing Repayment Effort of Self-Admitted Technical Debt

ResearchDGX agent

arXiv:2309.06020v3 Announce Type: replace-cross Abstract: Technical debt refers to the consequences of sub-optimal decisions made during software development that prioritize short-term benefits over l

The Alignment Target Problem: Divergent Moral Judgments of Humans, AI Systems, and Their Designers

Model ReleasesDGX agent

arXiv:2604.24155v1 Announce Type: cross Abstract: The quest to align machine behavior with human values raises fundamental questions about the moral frameworks that should govern AI decision-making. M

The Consensus Trap: Dissecting Subjectivity and the 'Ground Truth' Illusion in Data Annotation

SafetyDGX agent

arXiv:2602.11318v3 Announce Type: replace Abstract: In machine learning, 'ground truth' refers to the assumed correct labels used to train and evaluate models. However, the foundational 'ground truth'

The High Cost of Incivility: Quantifying Interaction Inefficiency via Multi-Agent Monte Carlo Simulations

AgentsDGX agent

arXiv:2512.08345v2 Announce Type: replace Abstract: Workplace toxicity is widely recognized as detrimental to organizational culture, yet quantifying its direct impact on operational efficiency remain

The Imbalanced User-AI Relationships as an Ethical Failure of Front-End Design in Healthcare AI

SafetyDGX agent

arXiv:2604.22767v1 Announce Type: cross Abstract: Ethical discourse on AI in healthcare has focused predominantly on back-end concerns such as bias, fairness and explainability, while the front-end in

The Kerimov-Alekberli Model: An Information-Geometric Framework for Real-Time System Stability

Model ReleasesDGX agent

arXiv:2604.24083v1 Announce Type: new Abstract: This study introduces the Kerimov-Alekberli model, a novel information-geometric framework that redefines AI safety by formally linking non-equilibrium

The Override Gap: A Magnitude Account of Knowledge Conflict Failure in Hypernetwork-Based Instant LLM Adaptation

Model ReleasesDGX agent

arXiv:2604.23750v1 Announce Type: cross Abstract: Hypernetwork-based methods such as Doc-to-LoRA internalize a document into an LLM's weights in a single forward pass, but they fail systematically on

The Power of Power Law: Asymmetry Enables Compositional Reasoning

TutorialsDGX agent

arXiv:2604.22951v1 Announce Type: new Abstract: Natural language data follows a power-law distribution, with most knowledge and skills appearing at very low frequency. While a common intuition suggest

The Pragmatic Persona: Discovering LLM Persona through Bridging Inference

Model ReleasesDGX agent

arXiv:2604.24079v1 Announce Type: cross Abstract: Large Language Models (LLMs) reveal inherent and distinctive personas through dialogue. However, most existing persona discovery approaches rely on su

The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications

Model ReleasesDGX agent

arXiv:2604.24668v1 Announce Type: new Abstract: Given the increased use of LLMs in financial systems today, it becomes important to evaluate the safety and robustness of such systems. One failure mode

The Randomness Floor: Measuring Intrinsic Non-Randomness in Language Model Token Distributions

Model ReleasesDGX agent

arXiv:2604.22771v1 Announce Type: cross Abstract: Language models cannot be random. This paper introduces Entropic Deviation (ED), the normalised KL divergence between a model's token distribution and

The Rise of Large Language Models and the Direction and Impact of US Federal Research Funding

Model ReleasesDGX agent

arXiv:2601.15485v2 Announce Type: replace-cross Abstract: Federal research funding shapes the direction, diversity, and impact of the US scientific enterprise. Large language models (LLMs) are rapidly

The Security Cost of Intelligence: AI Capability, Cyber Risk, and Deployment Paradox

Model ReleasesDGX agent

arXiv:2604.23058v1 Announce Type: cross Abstract: Firms are deploying more capable AI systems, but organizational controls often have not kept pace. These systems can generate greater productivity gai

The Shape of Attraction in UMAP: Exploring the Embedding Forces in Dimensionality Reduction

ResearchDGX agent

arXiv:2503.09101v4 Announce Type: replace-cross Abstract: Uniform manifold approximation and projection (UMAP) is among the most popular neighbor embedding methods. The method samples pairs of point i

The Spectral Lifecycle of Transformer Training: Transient Compression Waves, Persistent Spectral Gradients, and the Q/K--V Asymmetry

ResearchDGX agent

arXiv:2604.22778v1 Announce Type: cross Abstract: We present the first systematic study of weight matrix singular value spectra during transformer pretraining, tracking full SVD decompositions of ever

Think-at-Hard: Selective Latent Iterations to Improve Reasoning Language Models

Model ReleasesDGX agent

arXiv:2511.08577v2 Announce Type: replace-cross Abstract: Improving reasoning abilities of Large Language Models (LLMs), especially under parameter constraints, is crucial for real-world applications.

Thinking Like a Clinician: A Cognitive AI Agent for Clinical Diagnosis via Panoramic Profiling and Adversarial Debate

AgentsDGX agent

arXiv:2604.23605v1 Announce Type: new Abstract: The application of large language models (LLMs) in clinical decision support faces significant challenges of 'tunnel vision' and diagnostic hallucinatio

Time-Series Forecasting in Safety-Critical Environments: An EU-AI-Act-Compliant Open-Source Package / Zeitreihenprognose in sicherheitskritischen Umgebungen: Ein KI-VO-konformes Open-Source-Paket

SafetyDGX agent

arXiv:2604.23859v1 Announce Type: new Abstract: With spotforecast2-safe we present an integrated Compliance-by-Design approach to Python-based point forecasting of time series in safety-critical envir

Token Is All You Price

ApplicationsDGX agent

arXiv:2510.09859v4 Announce Type: replace-cross Abstract: A seller of a dynamic information service under an information-throughput constraint screens buyers who privately differ in urgency. We charac

Toward Polymorphic Backdoor against Semantic Communication via Intensity-Based Poisoning

Model ReleasesDGX agent

arXiv:2604.23231v1 Announce Type: cross Abstract: Semantic Communication (SC) backdoor attacks aim to utilize triggers to manipulate the system into producing predetermined outputs via backdoored shar

← Previous
1…302303304305306…354
Next →