AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
18 May 2026

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control

Model ReleasesDGX agent

arXiv:2605.15963v1 Announce Type: new Abstract: Large vision-language models have significantly advanced GUI agents, enabling executable interaction across web, mobile, and desktop interfaces. Yet the

Painless Activation Steering: An Automated, Lightweight Approach for Post-Training Large Language Models

Model ReleasesDGX agent

arXiv:2509.22739v3 Announce Type: replace-cross Abstract: Language models (LMs) are typically post-trained for desired capabilities and behaviors via weight-based or prompt-based steering, but the for

PanoWorld: Geometry-Consistent Panoramic Video World Modeling

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.15391v1 Announce Type: cross Abstract: We present PanoWorld, a panoramic video world model that generates geometry-consistent 360egree video from a single image and a caption. Existing pano

paper.json: A Coordination Convention for LLM-Agent-Actionable Papers

AgentsDGX agent

arXiv:2605.16194v1 Announce Type: cross Abstract: LLM agents routinely serve as first (and sometimes only) readers of academic papers, skimming for sub-claims, extracting reproducibility steps, and ge

PBT-Bench: Benchmarking AI Agents on Property-Based Testing

Model ReleasesDGX agent

arXiv:2605.15229v1 Announce Type: cross Abstract: Existing code benchmarks measure whether an agent can produce any test that reproduces a known bug, or whether it can produce a patch that fixes a des

PDRNN: Modular Data-driven Pedestrian Dead Reckoning on Loosely Coupled Radio- and Inertial-Signalstreams

Model ReleasesDGX agent

arXiv:2605.15252v1 Announce Type: cross Abstract: Modern pedestrian dead reckoning (PDR) systems rely on fusing noisy and biased estimates of position, velocity, and calibrated orientation derived fro

Petri Net Induced Heuristic Search for Resource Constrained Scheduling

ResearchDGX agent

arXiv:2605.15983v1 Announce Type: new Abstract: We formulate the Resource-Constrained Project Scheduling Problem (RCPSP) as optimal search over the reachability graph of a Timed Transition Petri Net w

PhysBrain 1.0 Technical Report

ResearchDGX agent

arXiv:2605.15298v1 Announce Type: cross Abstract: Vision-language-action models have advanced rapidly, but robot trajectories alone provide limited coverage for learning broad physical understanding.

Polynomial Neural Sheaf Diffusion: A Spectral Filtering Approach on Cellular Sheaves

SafetyDGX agent

arXiv:2512.00242v3 Announce Type: replace-cross Abstract: Sheaf Neural Networks equip graph structures with a cellular sheaf: a geometric structure which assigns local vector spaces (stalks) and a lin

Position: Artificial Intelligence Needs Meta Intelligence -- the Case for Metacognitive AI

ApplicationsDGX agent

arXiv:2605.15567v1 Announce Type: new Abstract: This position paper argues for metacognition as a general design principle for creating more accurate, secure, and efficient AI. The metacognitive solut

Position: Early-Stage Quality Assurance in Annotation Pipelines Is More Cost-Effective Than Late-Stage Validation

Model ReleasesDGX agent

arXiv:2605.15714v1 Announce Type: cross Abstract: This position paper argues that the machine learning community should prioritize early-stage quality assurance in annotation pipelines over the prevai

Pretraining Objective Matters in Extreme Low-Data FGVC: A Backbone-Controlled Study

ResearchDGX agent

arXiv:2605.15599v1 Announce Type: cross Abstract: Extreme low-data fine-grained classification is common in expert domains where labeling is expensive, yet practitioners still need principled guidance

PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI

AgentsDGX agent

arXiv:2605.15665v1 Announce Type: new Abstract: Deploying large language model (LLM)-driven conversational agents in enterprise settings requires prompts that are simultaneously correct at launch and

PrismQuant: Rate-Distortion-Optimal Vector Quantization for Gaussian-Mixture Sources

Local AiDGX agent

arXiv:2605.15507v1 Announce Type: cross Abstract: For a Gaussian source under mean-squared error (MSE), classical transform coding is rate--distortion (RD) optimal: the Karhunen--Loeve transform (KLT)

Process Rewards with Learned Reliability

ResearchDGX agent

arXiv:2605.15529v1 Announce Type: cross Abstract: Process Reward Models (PRMs) provide step-level feedback for reasoning, but current PRMs usually output only a single reward score for each step. Down

Property-Guided LLM Program Synthesis for Planning

TutorialsDGX agent

arXiv:2605.16142v1 Announce Type: new Abstract: LLMs have shown impressive success in program synthesis, discovering programs that surpass prior solutions. However, these approaches rely on simple num

Prospective multi-pathogen disease forecasting using autonomous LLM-guided tree search

AgentsDGX agent

arXiv:2605.16238v1 Announce Type: new Abstract: Probabilistic forecasting of infectious diseases is crucial for public health but relies on labor-intensive manual model curation by expert modeling tea

Quantization Undoes Alignment: Bias Emergence in Compressed LLMs Across Models and Precision Levels

Model ReleasesDGX agent

arXiv:2605.15208v1 Announce Type: cross Abstract: Large Language Models are routinely compressed via post-training quantization to reduce inference costs and memory footprint for cloud and edge deploy

Quantum Artificial Intelligence for Mission-Critical Systems: Foundations, Architectural Elements, and Future Directions

SafetyDGX agent

arXiv:2511.09884v2 Announce Type: replace Abstract: Mission critical (MC) applications such as defense operations, energy management, cybersecurity, and aerospace control require reliable, determinist

RaPD: Resolution-Agnostic Pixel Diffusion via Semantics-Enriched Implicit Representations

ResearchDGX agent

arXiv:2605.15908v1 Announce Type: cross Abstract: Natural images are continuous, yet most generative models synthesize them on discrete grids, limiting resolution-flexible generation. Continuous neura

RAR: Retrieving And Ranking Augmented MLLMs for Visual Recognition

Model ReleasesDGX agent

arXiv:2403.13805v2 Announce Type: replace-cross Abstract: CLIP (Contrastive Language-Image Pre-training) uses contrastive learning from noise image-text pairs to excel at recognizing a wide array of c

Reading the Cell, Designing the Cure: Perturbation-Conditioned Molecular Diffusion for Function-Oriented Drug Design

ResearchDGX agent

arXiv:2605.15243v1 Announce Type: cross Abstract: When reliable target structures are unavailable at scale or phenotypes arise from dysregulated pathways, transcriptomic perturbations provide a system

Reasoners or Translators? Contamination-aware Evaluation and Neuro-Symbolic Robustness in Tax Law

ApplicationsDGX agent

arXiv:2605.16052v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have significantly enhanced automated legal reasoning. Yet, it remains unclear whether their performance

RecMem: Recurrence-based Memory Consolidation for Efficient and Effective Long-Running LLM Agents

AgentsDGX agent

arXiv:2605.16045v1 Announce Type: cross Abstract: Memory systems often organize user-agent interactions as retrievable external memory and are crucial for long-running agents by overcoming the limited

Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective

SafetyDGX agent

arXiv:2605.15976v1 Announce Type: cross Abstract: Production machine translation relies overwhelmingly on encoder-decoder Seq2Seq models, yet reinforcement learning approaches to MT fine-tuning have l

Representation Without Reward: A JEPA Audit for LLM Fine-Tuning

Model ReleasesDGX agent

arXiv:2605.15394v1 Announce Type: cross Abstract: Joint-embedding predictive architectures (JEPAs) propose that a model should learn more useful abstractions when trained to predict latent representat

Residual Reinforcement Learning for Robot Teleoperation under Stochastic Delays

SafetyDGX agent

arXiv:2605.15480v1 Announce Type: cross Abstract: Stochastic communication delays in teleoperation introduce signal discontinuities that undermine control stability and degrade control performance. Co

Retrieval-Augmented Large Language Models for Schema-Constrained Clinical Information Extraction

Model ReleasesDGX agent

arXiv:2605.15467v1 Announce Type: cross Abstract: Conversational nurse-patient transcripts contain actionable observations, but converting these transcripts into structured representations at scale re

RIDE: Retinex-Informed Decoupling for Exposing Concealed Objects

ResearchDGX agent

arXiv:2605.15450v1 Announce Type: cross Abstract: Concealed Object Segmentation (COS) encompasses a family of dense-prediction tasks, including camouflaged object detection, polyp segmentation, transp

RoadmapBench: Evaluating Long-Horizon Agentic Software Development Across Version Upgrades

Model ReleasesDGX agent

arXiv:2605.15846v1 Announce Type: cross Abstract: Coding agents are increasingly deployed in real software development, where a single version iteration requires months of coordinated work across many

Robust Prior-Guided Segmentation for Editable 3D Gaussian Splatting

ResearchDGX agent

arXiv:2605.16065v1 Announce Type: cross Abstract: 3D Gaussian Splatting (3D-GS) enables real-time 3D scene reconstruction but lacks robust segmentation for editing tasks such as object removal, extrac

RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably

SafetyDGX agent

arXiv:2605.15514v1 Announce Type: cross Abstract: We identify intrinsic limitations of Rotary Positional Embeddings (RoPE) in Transformer-based long-context language models. Our theoretical analysis a

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

Model ReleasesDGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

Runtime-Structured Task Decomposition for Agentic Coding Systems

AgentsDGX agent

arXiv:2605.15425v1 Announce Type: cross Abstract: Agentic coding systems increasingly use large language models (LLMs) for software engineering tasks such as debugging, root cause analysis, and code r

SaaS-Bench: Can Computer-Use Agents Leverage Real-World SaaS to Solve Professional Workflows?

Model ReleasesDGX agent

arXiv:2605.15777v1 Announce Type: new Abstract: Computer-Using Agents (CUAs) are rapidly extending large language models (LLMs) beyond text-based reasoning toward action execution in more complex envi

SAE-RNA: A Sparse Autoencoder Model for Interpreting RNA Language Model Representations

ResearchDGX agent

arXiv:2510.02734v2 Announce Type: replace-cross Abstract: Deep learning, particularly with the advancement of Large Language Models, has transformed biomolecular modeling, with protein language models

SafeGPT: Preventing Data Leakage and Unethical Outputs in Enterprise LLM Use

SafetyDGX agent

arXiv:2601.06366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are transforming enterprise workflows but introduce security and ethics challenges when employees inadvertently s

SARVLM: A Vision Language Foundation Model for Semantic Understanding in SAR Imagery

Model ReleasesDGX agent

arXiv:2510.22665v3 Announce Type: replace-cross Abstract: Synthetic Aperture Radar (SAR) is a critical imaging modality due to its all-weather operational capability. Although recent advances in self-

ScreenSearch: Uncertainty-Aware OS Exploration

SafetyDGX agent

arXiv:2605.16024v1 Announce Type: new Abstract: Desktop GUI agents operate under partial observability: visually similar screens can correspond to different underlying workflow states, so locally plau

SDOF: Taming the Alignment Tax in Multi-Agent Orchestration with State-Constrained Dispatch

Model ReleasesDGX agent

arXiv:2605.15204v1 Announce Type: new Abstract: Multi-agent orchestration frameworks such as LangChain, LangGraph, and CrewAI route tasks through graph-based pipelines but do not enforce the stage con

Second-Order Multi-Level Variance Correction for Modality Competition in Multimodal Models

SafetyDGX agent

arXiv:2605.16165v1 Announce Type: cross Abstract: Autoregressive next-token training offers a unified formulation for image generation and text understanding, but it also creates strong modality compe

See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation

Local AiDGX agent

arXiv:2605.15585v1 Announce Type: new Abstract: Large language models can generate executable code for educational animations, but the resulting renders often exhibit visual defects, including element

Seeing is Understanding: Unlocking Causal Attention into Modality-Mutual Attention for Multimodal LLMs

ResearchDGX agent

arXiv:2503.02597v3 Announce Type: replace-cross Abstract: Recent Multimodal Large Language Models (MLLMs) have demonstrated significant progress in perceiving and reasoning over multimodal inquiries,

SemanticOpt: Towards LLM-Based Semantic Black-Box Optimization

Model ReleasesDGX agent

arXiv:2510.25404v3 Announce Type: replace-cross Abstract: Optimizing an experimental system can be extremely challenging when each experiment is expensive, time-consuming, or difficult to perform. Exi

Shaping Sparse Rewards in Reinforcement Learning: A Semi-supervised Approach

AgentsDGX agent

arXiv:2501.19128v5 Announce Type: replace-cross Abstract: In many real-world scenarios, reward signal for agents are exceedingly sparse, making it challenging to learn an effective reward function for

Shapley Neuron Values for Continual Learning: Which Neurons Matter Most?

TutorialsDGX agent

arXiv:2605.15877v1 Announce Type: cross Abstract: Continual learning enables neural networks to learn tasks sequentially without forgetting previously acquired knowledge. However, neural networks suff

Sharp Spectral Thresholds for Logit Fixed Points

ResearchDGX agent

arXiv:2605.15651v1 Announce Type: cross Abstract: Softmax feedback systems are a common mathematical core of entropy-regularized reinforcement learning, logit game dynamics, population choice, and mea

ShopGym: An Integrated Framework for Realistic Simulation and Scalable Benchmarking of E-Commerce Web Agents

Model ReleasesDGX agent

arXiv:2605.16116v1 Announce Type: new Abstract: Developing and evaluating e-commerce web agents requires environments that preserve meaningful task structure while enabling controllable, reproducible,

Sign-Separated Finite-Time Error Analysis of Q-Learning

SafetyDGX agent

arXiv:2605.16103v1 Announce Type: new Abstract: This paper develops a sign-separated finite-time error analysis for constant step-size Q-learning. Starting from the switching-system representation, th

SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces

Model ReleasesDGX agent

arXiv:2605.15215v1 Announce Type: new Abstract: Recently, skills have been widely adopted in large language model (LLM)-based agent systems across various domains. In existing frameworks, skills are t

SkiP: When to Skip and When to Refine for Efficient Robot Manipulation

SafetyDGX agent

arXiv:2605.15536v1 Announce Type: cross Abstract: Previous imitation learning policies predict future actions at every control step, whether in smooth motion phases or precise, contact-rich operation

SLIP & ETHICS: Graduated Intervention for AI Emotional Companions

SafetyDGX agent

arXiv:2605.15915v1 Announce Type: cross Abstract: AI emotional companions face a safety-rapport paradox: restrictive safeguards can damage supportive alliance, while permissive systems risk user harm.

Small Generalizable Prompt Predictive Models Can Steer Efficient RL Post-Training of Large Reasoning Models

ResearchDGX agent

arXiv:2602.01970v2 Announce Type: replace Abstract: Reinforcement learning enhances the reasoning capabilities of large language models but often involves high computational costs due to rollout-inten

SMCEvolve: Principled Scientific Discovery via Sequential Monte Carlo Evolution

TutorialsDGX agent

arXiv:2605.15308v1 Announce Type: new Abstract: LLM-driven program evolution has emerged as a powerful tool for automated scientific discovery, yet existing frameworks offer no principled guide for de

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

AgentsDGX agent

arXiv:2605.15301v1 Announce Type: new Abstract: Large language models (LLMs) still struggle with the rigorous reasoning demands of hard competitive programming. While recent multi-agent frameworks att

STAR: A Stage-attributed Triage and Repair framework for RCA Agents in Microservices

Model ReleasesDGX agent

arXiv:2605.15581v1 Announce Type: new Abstract: LLM-based root cause analysis (RCA) agents have recently emerged as a promising paradigm for incident diagnosis in microservice AIOps. However, their re

Stock Market Prediction Using Node Transformer Architecture Integrated with BERT Sentiment Analysis

ResearchDGX agent

arXiv:2603.05917v3 Announce Type: replace-cross Abstract: Stock market prediction presents considerable challenges for investors, financial institutions, and policymakers operating in complex market e

Structure Abstraction and Generalization in a Hippocampal-Entorhinal Inspired World Model

Model ReleasesDGX agent

arXiv:2605.15733v1 Announce Type: cross Abstract: Humans abstract experiences into structured representations to facilitate pattern inference and knowledge transfer. While the hippocampal-entorhinal (

Structure-BiEval: A Self-Supervised, Dual-Track Framework for Decoupling Structure and Content in LLM Evaluation for Web Information Systems

Model ReleasesDGX agent

arXiv:2601.19923v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) evolve into the core of Web-based autonomous agents and complex Web Information Systems, their ability to fait

Sufficient Explanations in Databases and their Connections to Database Repairs

TutorialsDGX agent

arXiv:2511.15623v2 Announce Type: replace-cross Abstract: We investigate the notion of sufficient explanation, and a sufficiency-degree as attribution score for database tuples in relation to query an

← Previous
1…248249250251252…358
Next →