AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
13 Apr 2026

Can We Still Hear the Accent? Investigating the Resilience of Native Language Signals in the LLM Era

ResearchDGX agent

arXiv:2604.08568v1 Announce Type: cross Abstract: The evolution of writing assistance tools from machine translation to large language models (LLMs) has changed how researchers write. This study inves

Cards Against LLMs: Benchmarking Humor Alignment in Large Language Models

SafetyDGX agent

arXiv:2604.08757v1 Announce Type: cross Abstract: Humor is one of the most culturally embedded and socially significant dimensions of human communication, yet it remains largely unexplored as a dimens

Case-Grounded Evidence Verification: A Framework for Constructing Evidence-Sensitive Supervision

Local AiDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.09537v1 Announce Type: cross Abstract: Evidence-grounded reasoning requires more than attaching retrieved text to a prediction: a model should make decisions that depend on whether the prov

Chain-in-Tree: Back to Sequential Reasoning in LLM Tree Search

SafetyDGX agent

arXiv:2509.25835v4 Announce Type: replace Abstract: Test-time scaling improves large language models (LLMs) on long-horizon reasoning tasks by allocating more compute at inference. LLM inference via t

Chain-of-Zoom: Extreme Super-Resolution via Scale Autoregression and Preference Alignment

SafetyDGX agent

arXiv:2505.18600v3 Announce Type: replace-cross Abstract: Modern single-image super-resolution (SISR) models deliver photo-realistic results at the scale factors on which they are trained, but collaps

ChipSeek: Optimizing Verilog Generation via EDA-Integrated Reinforcement Learning

SafetyDGX agent

arXiv:2507.04736v2 Announce Type: replace Abstract: Large Language Models have emerged as powerful tools for automating Register-Transfer Level (RTL) code generation, yet they face critical limitation

Chronological Contrastive Learning: Few-Shot Progression Assessment in Irreversible Diseases

ResearchDGX agent

arXiv:2603.21935v2 Announce Type: replace-cross Abstract: Quantitative disease severity scoring in medical imaging is costly, time-consuming, and subject to inter-reader variability. At the same time,

CLIP-Inspector: Model-Level Backdoor Detection for Prompt-Tuned CLIP via OOD Trigger Inversion

ResearchDGX agent

arXiv:2604.09101v1 Announce Type: cross Abstract: Organisations with limited data and computational resources increasingly outsource model training to Machine Learning as a Service (MLaaS) providers,

Commanding Humanoid by Free-form Language: A Large Language Action Model with Unified Motion Vocabulary

SafetyDGX agent

arXiv:2511.22963v2 Announce Type: replace-cross Abstract: Enabling humanoid robots to follow free-form language commands is critical for seamless human-robot interaction, collaborative task execution,

CONDESION-BENCH: Conditional Decision-Making of Large Language Models in Compositional Action Space

Model ReleasesDGX agent

arXiv:2604.09029v1 Announce Type: cross Abstract: Large language models have been widely explored as decision-support tools in high-stakes domains due to their contextual understanding and reasoning c

Constraining Sequential Model Editing with Editing Anchor Compression

Model ReleasesDGX agent

arXiv:2503.00035v2 Announce Type: replace-cross Abstract: Large language models (LLMs) struggle with hallucinations due to false or outdated knowledge. Given the high resource demands of retraining th

Constraint-Aware Corrective Memory for Language-Based Drug Discovery Agents

SafetyDGX agent

arXiv:2604.09308v1 Announce Type: new Abstract: Large language models are making autonomous drug discovery agents increasingly feasible, but reliable success in this setting is not determined by any s

CORA: Conformal Risk-Controlled Agents for Safeguarded Mobile GUI Automation

Model ReleasesDGX agent

arXiv:2604.09155v1 Announce Type: cross Abstract: Graphical user interface (GUI) agents powered by vision language models (VLMs) are rapidly moving from passive assistance to autonomous operation. How

CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference

HardwareDGX agent

arXiv:2604.08584v1 Announce Type: cross Abstract: Long-context LLMs increasingly rely on extended, reusable prefill prompts for agents and domain Q&A, pushing attention and KV-cache to become the domi

DDSP-QbE++: Improving Speech Quality for Speech Anonymisation for Atypical Speech

ResearchDGX agent

arXiv:2604.09246v1 Announce Type: cross Abstract: Differentiable Digital Signal Processing (DDSP) pipelines for voice conversion rely on subtractive synthesis, where a periodic excitation signal is sh

Decomposing the Delta: What Do Models Actually Learn from Preference Pairs?

TutorialsDGX agent

arXiv:2604.08723v1 Announce Type: cross Abstract: Preference optimization methods such as DPO and KTO are widely used for aligning language models, yet little is understood about what properties of pr

Deep Learning-Based Tracking and Lineage Reconstruction of Ligament Breakup

ResearchDGX agent

arXiv:2604.08711v1 Announce Type: cross Abstract: The disintegration of liquid sheets into ligaments and droplets involves highly transient, multi-scale dynamics that are difficult to quantify from hi

DeepGuard: Secure Code Generation via Multi-Layer Semantic Aggregation

ResearchDGX agent

arXiv:2604.09089v1 Announce Type: cross Abstract: Large Language Models (LLMs) for code generation can replicate insecure patterns from their training data. To mitigate this, a common strategy for sec

Dejavu: Towards Experience Feedback Learning for Embodied Intelligence

SafetyDGX agent

arXiv:2510.10181v3 Announce Type: replace-cross Abstract: Embodied agents face a fundamental limitation: once deployed in real-world environments, they cannot easily acquire new knowledge to improve t

Demystifying the Silence of Correctness Bugs in PyTorch Compiler

ResearchDGX agent

arXiv:2604.08720v1 Announce Type: cross Abstract: Performance optimization of AI infrastructure is key to the fast adoption of large language models (LLMs). The PyTorch compiler (torch.compile), a cor

Descriptor: Parasitoid Wasps and Associated Hymenoptera Dataset (DAPWH)

ResearchDGX agent

arXiv:2602.20028v2 Announce Type: replace-cross Abstract: Accurate taxonomic identification is the cornerstone of biodiversity monitoring and agricultural management, particularly for the hyper-divers

Detection and Characterization of Coordinated Online Behavior: A Survey

TutorialsDGX agent

arXiv:2408.01257v2 Announce Type: replace-cross Abstract: Coordination is a fundamental aspect of life. The advent of social media has made it integral also to online human interactions, such as those

Detection of Hate and Threat in Digital Forensics: A Case-Driven Multimodal Approach

ResearchDGX agent

arXiv:2604.08609v1 Announce Type: cross Abstract: Digital forensic investigations increasingly rely on heterogeneous evidence such as images, scanned documents, and contextual reports. These artifacts

Dictionary-Aligned Concept Control for Safeguarding Multimodal LLMs

SafetyDGX agent

arXiv:2604.08846v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have been shown to be vulnerable to malicious queries that can elicit unsafe responses. Recent work uses prom

DiffHDR: Re-Exposing LDR Videos with Video Diffusion Models

ApplicationsDGX agent

arXiv:2604.06161v2 Announce Type: replace-cross Abstract: Most digital videos are stored in 8-bit low dynamic range (LDR) formats, where much of the original high dynamic range (HDR) scene radiance is

Distilling Genomic Models for Efficient mRNA Representation Learning via Embedding Matching

ResearchDGX agent

arXiv:2604.08574v1 Announce Type: cross Abstract: Large Genomic Foundation Models have recently achieved remarkable results and in-vivo translation capabilities. However these models quickly grow to o

Distributionally Robust Token Optimization in RLHF

ResearchDGX agent

arXiv:2604.08577v1 Announce Type: cross Abstract: Large Language Models (LLMs) tend to respond correctly to prompts that align to the data they were trained and fine-tuned on. Yet, small shifts in wor

Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies

SafetyDGX agent

arXiv:2604.09189v1 Announce Type: cross Abstract: LLMs internalize safety policies through RLHF, yet these policies are never formally specified and remain difficult to inspect. Existing benchmarks ev

Do We Really Need to Approach the Entire Pareto Front in Many-Objective Bayesian Optimisation?

Model ReleasesDGX agent

arXiv:2604.09417v1 Announce Type: new Abstract: Many-objective optimisation, a subset of multi-objective optimisation, involves optimisation problems with more than three objectives. As the number of

DRBENCHER: Can Your Agent Identify the Entity, Retrieve Its Properties and Do the Math?

Model ReleasesDGX agent

arXiv:2604.09251v1 Announce Type: new Abstract: Deep research agents increasingly interleave web browsing with multi-step computation, yet existing benchmarks evaluate these capabilities in isolation,

Drift and selection in LLM text ecosystems

TutorialsDGX agent

arXiv:2604.08554v1 Announce Type: cross Abstract: The public text record -- the material from which both people and AI systems now learn -- is increasingly shaped by its own outputs. Generated text en

Dynamic sparsity in tree-structured feed-forward layers at scale

ResearchDGX agent

arXiv:2604.08565v1 Announce Type: cross Abstract: At typical context lengths, the feed-forward MLP block accounts for a large share of a transformer's compute budget, motivating sparse alternatives to

E3-TIR: Enhanced Experience Exploitation for Tool-Integrated Reasoning

SafetyDGX agent

arXiv:2604.09455v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated significant potential in Tool-Integrated Reasoning (TIR), existing training paradigms face signific

eBandit: Kernel-Driven Reinforcement Learning for Adaptive Video Streaming

ApplicationsDGX agent

arXiv:2604.08791v1 Announce Type: cross Abstract: User-space Adaptive Bitrate (ABR) algorithms cannot see the transport layer signals that matter most, such as minimum RTT and instantaneous delivery r

ECHO: Efficient Chest X-ray Report Generation with One-step Block Diffusion

SafetyDGX agent

arXiv:2604.09450v1 Announce Type: cross Abstract: Chest X-ray report generation (CXR-RG) has the potential to substantially alleviate radiologists' workload. However, conventional autoregressive visio

EchoTrail-GUI: Building Actionable Memory for GUI Agents via Critic-Guided Self-Exploration

AgentsDGX agent

arXiv:2512.19396v3 Announce Type: replace Abstract: Contemporary GUI agents, while increasingly capable due to advances in Large Vision-Language Models (VLMs), often operate with a critical limitation

EGMOF: Efficient Generation of Metal-Organic Frameworks Using a Hybrid Diffusion-Transformer Architecture

ResearchDGX agent

arXiv:2511.03122v2 Announce Type: replace-cross Abstract: Designing materials with targeted properties remains challenging due to the vastness of chemical space and the scarcity of property-labeled da

EigentSearch-Q+: Enhancing Deep Research Agents with Structured Reasoning Tools

Model ReleasesDGX agent

arXiv:2604.07927v2 Announce Type: replace Abstract: Deep research requires reasoning over web evidence to answer open-ended questions, and it is a core capability for AI agents. Yet many deep research

EMA Is Not All You Need: Mapping the Boundary Between Structure and Content in Recurrent Context

Model ReleasesDGX agent

arXiv:2604.08556v1 Announce Type: cross Abstract: What exactly do efficient sequence models gain over simple temporal averaging? We use exponential moving average (EMA) traces, the simplest recurrent

Enhancing LLM Problem Solving via Tutor-Student Multi-Agent Interaction

Model ReleasesDGX agent

arXiv:2604.08931v1 Announce Type: new Abstract: Human cognitive development is shaped not only by individual effort but by structured social interaction, where role-based exchanges such as those betwe

Enhancing the Safety of Medical Vision-Language Models by Synthetic Demonstrations

SafetyDGX agent

arXiv:2506.09067v2 Announce Type: replace-cross Abstract: Generative medical vision-language models~(Med-VLMs) are primarily designed to generate complex textual information~(e.g., diagnostic reports)

Envisioning the Future, One Step at a Time

Model ReleasesDGX agent

arXiv:2604.09527v1 Announce Type: cross Abstract: Accurately anticipating how complex, diverse scenes will evolve requires models that represent uncertainty, simulate along extended interaction chains

EquiformerV3: Scaling Efficient, Expressive, and General SE(3)-Equivariant Graph Attention Transformers

ResearchDGX agent

arXiv:2604.09130v1 Announce Type: cross Abstract: As SE(3)-equivariant graph neural networks mature as a core tool for 3D atomistic modeling, improving their efficiency, expressivity, and physical con

Every Response Counts: Quantifying Uncertainty of LLM-based Multi-Agent Systems through Tensor Decomposition

AgentsDGX agent

arXiv:2604.08708v1 Announce Type: cross Abstract: While Large Language Model-based Multi-Agent Systems (MAS) consistently outperform single-agent systems on complex tasks, their intricate interactions

Evidential Transformation Network: Turning Pretrained Models into Evidential Models for Post-hoc Uncertainty Estimation

ApplicationsDGX agent

arXiv:2604.08627v1 Announce Type: cross Abstract: Pretrained models have become standard in both vision and language, yet they typically do not provide reliable measures of confidence. Existing uncert

Evolutionary Optimization Trumps Adam Optimization on Embedding Space Exploration

SafetyDGX agent

arXiv:2511.03913v2 Announce Type: replace-cross Abstract: Deep diffusion models have revolutionized image generation by producing high-quality outputs. However, achieving specific objectives with thes

Explorable Theorems: Making Written Theorems Explorable by Grounding Them in Formal Representations

ResearchDGX agent

arXiv:2604.02598v2 Announce Type: replace-cross Abstract: LLM-generated explanations can make technical content more accessible, but there is a ceiling on what they can support interactively. Because

Exploring Teachers' Perspectives on Using Conversational AI Agents for Group Collaboration

SafetyDGX agent

arXiv:2602.07142v2 Announce Type: replace-cross Abstract: Collaboration is a cornerstone of 21st-century learning, yet teachers continue to face challenges in supporting productive peer interaction. E

Extrapolating Volition with Recursive Information Markets

SafetyDGX agent

arXiv:2604.08606v1 Announce Type: cross Abstract: One of the impediments to the efficiency of information markets is the inherent information asymmetry present in them, exacerbated by the 'buyer's ins

Fine-tuning is Not Enough: A Parallel Framework for Collaborative Imitation and Reinforcement Learning in End-to-end Autonomous Driving

Model ReleasesDGX agent

arXiv:2603.13842v3 Announce Type: replace-cross Abstract: End-to-end autonomous driving is typically built upon imitation learning (IL), yet its performance is constrained by the quality of human demo

FluidFlow: a flow-matching generative model for fluid dynamics surrogates on unstructured meshes

Model ReleasesDGX agent

arXiv:2604.08586v1 Announce Type: cross Abstract: Computational fluid dynamics (CFD) provides high-fidelity simulations of fluid flows but remains computationally expensive for many-query applications

Frequency-Enhanced Diffusion Models: Curriculum-Guided Semantic Alignment for Zero-Shot Skeleton Action Recognition

SafetyDGX agent

arXiv:2604.09063v1 Announce Type: cross Abstract: Human action recognition is pivotal in computer vision, with applications ranging from surveillance to human-robot interaction. Despite the effectiven

From Business Events to Auditable Decisions: Ontology-Governed Graph Simulation for Enterprise AI

Model ReleasesDGX agent

arXiv:2604.08603v1 Announce Type: new Abstract: Existing LLM-based agent systems share a common architectural failure: they answer from the unrestricted knowledge space without first simulating how ac

From Dispersion to Attraction: Spectral Dynamics of Hallucination Across Whisper Model Scales

SafetyDGX agent

arXiv:2604.08591v1 Announce Type: cross Abstract: Hallucinations in large ASR models present a critical safety risk. In this work, we propose the extit{Spectral Sensitivity Theorem}, which predicts a

From Navigation to Refinement: Revealing the Two-Stage Nature of Flow-based Diffusion Models through Oracle Velocity

ResearchDGX agent

arXiv:2512.02826v3 Announce Type: replace-cross Abstract: Flow-based diffusion models have emerged as a leading paradigm for training generative models across images and videos. However, their memoriz

From Paper to Program: Accelerating Quantum Many-Body Algorithm Development via a Multi-Stage LLM-Assisted Workflow

Model ReleasesDGX agent

arXiv:2604.04089v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate code rapidly but remain unreliable for scientific algorithms whose correctness depends on structural

From Selection to Scheduling: Federated Geometry-Aware Correction Makes Exemplar Replay Work Better under Continual Dynamic Heterogeneity

SafetyDGX agent

arXiv:2604.08617v1 Announce Type: cross Abstract: Exemplar replay has become an effective strategy for mitigating catastrophic forgetting in federated continual learning (FCL) by retaining representat

GAN-Enhanced Deep Reinforcement Learning for Semantic-Aware Resource Allocation in 6G Network Slicing

SafetyDGX agent

arXiv:2604.08576v1 Announce Type: cross Abstract: Sixth-generation (6G) wireless networks must support heterogeneous services: enhanced Mobile Broadband (eMBB) requiring 1 Tbps data rates, massive Mac

Ge^ext{2}mS-T: Multi-Dimensional Grouping for Ultra-High Energy Efficiency in Spiking Transformer

ResearchDGX agent

arXiv:2604.08894v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer superior energy efficiency over Artificial Neural Networks (ANNs). However, they encounter significant deficienci

Gen-n-Val: Agentic Image Data Generation and Validation

AgentsDGX agent

arXiv:2506.04676v2 Announce Type: replace-cross Abstract: The data scarcity, label noise, and long-tailed category imbalance remain important and unresolved challenges in many computer vision tasks, s

← Previous
1…339340341342343…350
Next →