AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Local Ai

KRONE: Scalable LLM-Augmented Log Anomaly Detection via Hierarchical Abstraction

DGX agent

arXiv:2602.07303v3 Announce Type: replace-cross Abstract: Log anomaly detection is crucial for uncovering system failures and security risks. Although logs originate from nested component executions w

local-aiarxiv-cs-ai
20 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

KWBench: Measuring Unprompted Problem Recognition in Knowledge Work

DGX agent

arXiv:2604.15760v1 Announce Type: new Abstract: We introduce the first version of KWBench (Knowledge Work Bench), a benchmark for unprompted problem recognition in large language models: can an LLM id

model-releasesarxiv-cs-ai
20 Apr 2026
Research

LACE: Lattice Attention for Cross-thread Exploration

DGX agent

arXiv:2604.15529v1 Announce Type: new Abstract: Current large language models reason in isolation. Although it is common to sample multiple reasoning paths in parallel, these trajectories do not inter

researcharxiv-cs-ai
20 Apr 2026
Safety

Language Models as Semantic Teachers: Post-Training Alignment for Medical Audio Understanding

DGX agent

arXiv:2512.04847v2 Announce Type: replace-cross Abstract: Pre-trained audio models excel at detecting acoustic patterns in auscultation sounds but often fail to grasp their clinical significance, limi

safetyarxiv-cs-ai
20 Apr 2026
Safety

Large Language Models for Market Research: A Data-augmentation Approach

DGX agent

arXiv:2412.19363v3 Announce Type: replace Abstract: Large Language Models (LLMs) have transformed artificial intelligence by excelling in complex natural language processing tasks. Their ability to ge

safetyarxiv-cs-ai
20 Apr 2026
Research

Learning to Reason with Insight for Informal Theorem Proving

DGX agent

arXiv:2604.16278v1 Announce Type: new Abstract: Although most of the automated theorem-proving approaches depend on formal proof systems, informal theorem proving can align better with large language

researcharxiv-cs-ai
20 Apr 2026
Research

Learning Uncertainty from Sequential Internal Dispersion in Large Language Models

DGX agent

arXiv:2604.15741v1 Announce Type: cross Abstract: Uncertainty estimation is a promising approach to detect hallucinations in large language models (LLMs). Recent approaches commonly depend on model in

researcharxiv-cs-ai
20 Apr 2026
Research

Lightweight Geometric Adaptation for Training Physics-Informed Neural Networks

DGX agent

arXiv:2604.15392v1 Announce Type: cross Abstract: Physics-Informed Neural Networks (PINNs) often suffer from slow convergence, training instability, and reduced accuracy on challenging partial differe

researcharxiv-cs-ai
20 Apr 2026
Model Releases

LinuxArena: A Control Setting for AI Agents in Live Production Software Environments

DGX agent

arXiv:2604.15384v1 Announce Type: cross Abstract: We introduce LinuxArena, a control setting in which agents operate directly on live, multi-service production environments. LinuxArena contains 20 env

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

LLM attribution analysis across different fine-tuning strategies and model scales for automated code compliance

DGX agent

arXiv:2604.15589v1 Announce Type: cross Abstract: Existing research on large language models (LLMs) for automated code compliance has primarily focused on performance, treating the models as black box

model-releasesarxiv-cs-ai
20 Apr 2026
Research

LLM Reasoning Is Latent, Not the Chain of Thought

DGX agent

arXiv:2604.15726v1 Announce Type: new Abstract: This position paper argues that large language model (LLM) reasoning should be studied as latent-state trajectory formation rather than as faithful surf

researcharxiv-cs-ai
20 Apr 2026
Research

LLMbench: A Comparative Close Reading Workbench for Large Language Models

DGX agent

arXiv:2604.15508v1 Announce Type: cross Abstract: LLMbench is a browser-based workbench for the comparative close reading of large language model (LLM) outputs. Where existing tools for LLM comparison

researcharxiv-cs-ai
20 Apr 2026
Research

Losses that Cook: Topological Optimal Transport for Structured Recipe Generation

DGX agent

arXiv:2601.02531v2 Announce Type: replace-cross Abstract: Cooking recipes are complex procedures that require not only a fluent and factual text, but also accurate timing, temperature, and procedural

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Mamba-SSM with LLM Reasoning for Feature Selection: Faithfulness-Aware Biomarker Discovery

DGX agent

arXiv:2604.14334v2 Announce Type: replace-cross Abstract: Gradient saliency from deep sequence models surfaces candidate biomarkers efficiently, but the resulting gene lists can be contaminated by tis

model-releasesarxiv-cs-ai
20 Apr 2026
Local Ai

MambaBack: Bridging Local Features and Global Contexts in Whole Slide Image Analysis

DGX agent

arXiv:2604.15729v1 Announce Type: cross Abstract: Whole Slide Image (WSI) analysis is pivotal in computational pathology, enabling cancer diagnosis by integrating morphological and architectural cues

local-aiarxiv-cs-ai
20 Apr 2026
Agents

MARCH: Multi-Agent Radiology Clinical Hierarchy for CT Report Generation

DGX agent

arXiv:2604.16175v1 Announce Type: new Abstract: Automated 3D radiology report generation often suffers from clinical hallucinations and a lack of the iterative verification found in human practice. Wh

agentsarxiv-cs-ai
20 Apr 2026
Research

Mechanisms of Prompt-Induced Hallucination in Vision-Language Models

DGX agent

arXiv:2601.05201v2 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) are highly capable, yet often hallucinate by favoring textual prompts over visual evidence. We study this

researcharxiv-cs-ai
20 Apr 2026
Model Releases

MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition

DGX agent

arXiv:2604.16009v1 Announce Type: new Abstract: Metacognition, the ability to monitor and regulate one's own reasoning, remains under-evaluated in AI benchmarking. We introduce MEDLEY-BENCH, a benchma

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

MFC-RFNet: A Multi-scale Guided Rectified Flow Network for Radar Sequence Prediction

DGX agent

arXiv:2601.03633v2 Announce Type: replace-cross Abstract: Accurate and high-resolution precipitation nowcasting from radar echo sequences is crucial for disaster mitigation and economic planning, yet

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

Mind DeepResearch Technical Report

DGX agent

arXiv:2604.14518v2 Announce Type: replace Abstract: We present Mind DeepResearch (MindDR), an efficient multi-agent deep research framework that achieves leading performance with only ~30B-parameter m

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Mind's Eye: A Benchmark of Visual Abstraction, Transformation and Composition for Multimodal LLMs

DGX agent

arXiv:2604.16054v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have achieved impressive progress on vision language benchmarks, yet their capacity for visual cognitive and

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Mitigating hallucinations and omissions in LLMs for invertible problems: An application to hardware logic design automation

DGX agent

arXiv:2512.03053v2 Announce Type: replace-cross Abstract: We show for invertible problems that transform data from a source domain (for example, Logic Condition Tables (LCTs)) to a destination domain

researcharxiv-cs-ai
20 Apr 2026
Model Releases

MM-Telco: Benchmarks and Multimodal Large Language Models for Telecom Applications

DGX agent

arXiv:2511.13131v2 Announce Type: replace Abstract: Large Language Models (LLMs) have emerged as powerful tools for automating complex reasoning and decision-making tasks. In telecommunications, they

model-releasesarxiv-cs-ai
20 Apr 2026
Applications

Modeling of ASD/TD Children's Behaviors in Interaction with a Virtual Social Robot During a Music Education Program Using Deep Neural Networks

DGX agent

arXiv:2604.15314v1 Announce Type: cross Abstract: This research aimed to develop an intelligent system to evaluate performance and extract behavioral models for children with ASD and neurotypical (TD)

applicationsarxiv-cs-ai
20 Apr 2026
Applications

MRGEN: A Conceptual Framework for LLM-Powered Mixed Reality Authoring Tools for Education

DGX agent

arXiv:2604.15341v1 Announce Type: cross Abstract: Mixed Reality (MR) offers immersive and multimodal opportunities for education but remains difficult for teachers to author without technical expertis

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

MTR-DuplexBench: Towards a Comprehensive Evaluation of Multi-Round Conversations for Full-Duplex Speech Language Models

DGX agent

arXiv:2511.10262v3 Announce Type: replace-cross Abstract: Full-Duplex Speech Language Models (FD-SLMs) enable real-time, overlapping conversational interactions, offering a more dynamic user experienc

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Multi-View Attention Multiple-Instance Learning Enhanced by LLM Reasoning for Cognitive Distortion Detection

DGX agent

arXiv:2509.17292v3 Announce Type: replace-cross Abstract: Cognitive distortions have been closely linked to mental health disorders, yet their automatic detection remains challenging due to contextual

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Natural gradient descent with momentum

DGX agent

arXiv:2604.15554v1 Announce Type: cross Abstract: We consider the problem of approximating a function by an element of a nonlinear manifold which admits a differentiable parametrization, typical examp

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Neuro-Symbolic ODE Discovery with Latent Grammar Flow

DGX agent

arXiv:2604.16232v1 Announce Type: cross Abstract: Understanding natural and engineered systems often relies on symbolic formulations, such as differential equations, which provide interpretability and

researcharxiv-cs-ai
20 Apr 2026
Research

NeuroLip: An Event-driven Spatiotemporal Learning Framework for Cross-Scene Lip-Motion-based Visual Speaker Recognition

DGX agent

arXiv:2604.15718v1 Announce Type: cross Abstract: Visual speaker recognition based on lip motion offers a silent, hands-free, and behavior-driven biometric solution that remains effective even when ac

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Neurosymbolic Repo-level Code Localization

DGX agent

arXiv:2604.16021v1 Announce Type: cross Abstract: Code localization is a cornerstone of autonomous software engineering. Recent advancements have achieved impressive performance on real-world issue be

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Noise Aggregation Analysis Driven by Small-Noise Injection: Efficient Membership Inference for Diffusion Models

DGX agent

arXiv:2510.21783v2 Announce Type: replace-cross Abstract: Diffusion models have demonstrated powerful performance in generating high-quality images. A typical example is text-to-image generator like S

researcharxiv-cs-ai
20 Apr 2026
Model Releases

OjaKV: Context-Aware Online Low-Rank KV Cache Compression

DGX agent

arXiv:2509.21623v2 Announce Type: replace-cross Abstract: The expanding long-context capabilities of large language models are constrained by a significant memory bottleneck: the key-value (KV) cache

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Opportunities and Challenges of Large Language Models for Low-Resource Languages in Humanities Research

DGX agent

arXiv:2412.04497v5 Announce Type: replace-cross Abstract: Low-resource languages serve as invaluable repositories of human history, embodying cultural evolution and intellectual diversity. Despite the

researcharxiv-cs-ai
20 Apr 2026
Model Releases

OSCBench: Benchmarking Object State Change in Text-to-Video Generation

DGX agent

arXiv:2603.11698v2 Announce Type: replace-cross Abstract: Text-to-video (T2V) generation models have made rapid progress in producing visually high-quality and temporally coherent videos. However, exi

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

PAWN: Piece Value Analysis with Neural Networks

DGX agent

arXiv:2604.15585v1 Announce Type: cross Abstract: Predicting the relative value of any given chess piece in a position remains an open challenge, as a piece's contribution depends on its spatial relat

safetyarxiv-cs-ai
20 Apr 2026
Safety

Persona-Assigned Large Language Models Exhibit Human-Like Motivated Reasoning

DGX agent

arXiv:2506.20020v2 Announce Type: replace Abstract: Reasoning in humans is prone to biases due to underlying motivations like identity protection, that undermine rational decision-making and judgment.

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

Phase Transitions as the Breakdown of Statistical Indistinguishability

DGX agent

arXiv:2604.15773v1 Announce Type: cross Abstract: We introduce a novel characterization of phase transitions based on hypothesis testing. In our formulation, a phase transition is defined as the break

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

PIIBench: A Unified Multi-Source Benchmark Corpus for Personally Identifiable Information Detection

DGX agent

arXiv:2604.15776v1 Announce Type: cross Abstract: We present PIIBench, a unified benchmark corpus for Personally Identifiable Information (PII) detection in natural language text. Existing resources f

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Polarization by Default: Auditing Recommendation Bias in LLM-Based Content Curation

DGX agent

arXiv:2604.15937v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed to curate and rank human-created content, yet the nature and structure of their biases in these

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

PolicyBank: Evolving Policy Understanding for LLM Agents

DGX agent

arXiv:2604.15505v1 Announce Type: cross Abstract: LLM agents operating under organizational policies must comply with authorization constraints typically specified in natural language. In practice, su

model-releasesarxiv-cs-ai
20 Apr 2026
Local Ai

Power to the Clients: Federated Learning in a Dictatorship Setting

DGX agent

arXiv:2510.22149v3 Announce Type: replace-cross Abstract: Federated learning (FL) has emerged as a promising paradigm for decentralized model training, enabling multiple clients to collaboratively lea

local-aiarxiv-cs-ai
20 Apr 2026
Safety

Preregistered Belief Revision Contracts

DGX agent

arXiv:2604.15558v1 Announce Type: new Abstract: Deliberative multi-agent systems allow agents to exchange messages and revise beliefs over time. While this interaction is meant to improve performance,

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

Prices, Bids, Values: One ML-Powered Combinatorial Auction to Rule Them All

DGX agent

arXiv:2411.09355v3 Announce Type: replace-cross Abstract: We study the design of iterative combinatorial auctions (ICAs). The main challenge in this domain is that the bundle space grows exponentially

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Privacy-Preserving LLMs Routing

DGX agent

arXiv:2604.15728v1 Announce Type: cross Abstract: Large language model (LLM) routing has emerged as a critical strategy to balance model performance and cost-efficiency by dynamically selecting servic

researcharxiv-cs-ai
20 Apr 2026
Model Releases

PRL-Bench: A Comprehensive Benchmark Evaluating LLMs' Capabilities in Frontier Physics Research

DGX agent

arXiv:2604.15411v1 Announce Type: cross Abstract: The paradigm of agentic science requires AI systems to conduct robust reasoning and engage in long-horizon, autonomous exploration. However, current s

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Protecting Language Models Against Unauthorized Distillation through Trace Rewriting

DGX agent

arXiv:2602.15143v2 Announce Type: replace Abstract: Knowledge distillation is a widely adopted technique for transferring capabilities from LLMs to smaller, more efficient student models. However, una

researcharxiv-cs-ai
20 Apr 2026
Safety

Prototype-Grounded Concept Models for Verifiable Concept Alignment

DGX agent

arXiv:2604.16076v1 Announce Type: cross Abstract: Concept Bottleneck Models (CBMs) aim to improve interpretability in Deep Learning by structuring predictions through human-understandable concepts, bu

safetyarxiv-cs-ai
20 Apr 2026
← Previous
1…403404405406407…443
Next →