AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
11 Jun 2026

Natural-Language Temporal Grounding in Hour-Long Videos is a Search Problem: A Benchmark and Empirical Decomposition

Model ReleasesDGX agent

arXiv:2606.12300v1 Announce Type: cross Abstract: Temporal grounding--returning the interval [t_s, t_e] for a natural-language query over a video--is the language interface to long-form video, yet has

nD-RoPE: A Generalized RoPE for n-Dimensional Position Embedding

ResearchDGX agent

arXiv:2606.12146v1 Announce Type: cross Abstract: Rotary Position Embedding (RoPE) is widely adopted in Transformer models, yet its extension to high-dimensional domains lacks a unified theoretical fo

Neural FOXP2 -- Language Specific Neuron Steering for Targeted Language Improvement in LLMs

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.00945v2 Announce Type: replace-cross Abstract: LLMs are multilingual by training, yet their lingua franca is often English, reflecting English language dominance in pretraining. Other langu

NightFeats @ MMU-RAGent NeurIPS 2025: A Context-Optimized Multi-Agent RAG System for the Text-to-Text Track

Model ReleasesDGX agent

arXiv:2606.11199v1 Announce Type: cross Abstract: We present NightFeats, a structured multi-agent retrieval-augmented generation (RAG) system submitted to the MMU-RAGent competition at NeurIPS 2025, w

Noise-Aware Framework for Correcting Corrupted Labels

ApplicationsDGX agent

arXiv:2606.11695v1 Announce Type: cross Abstract: High-quality labeled data is essential for training reliable ML/DL models. However, real-world datasets often contain a considerable proportion of cor

Noise-Guided Transport for Imitation Learning

SafetyDGX agent

arXiv:2509.26294v2 Announce Type: replace-cross Abstract: We consider imitation learning in the low-data regime, where only a limited number of expert demonstrations are available. In this setting, me

Non-frontal face recognition using GANs and memristor-based classifiers

Local AiDGX agent

arXiv:2606.12074v1 Announce Type: cross Abstract: Face recognition systems have advanced significantly through deep learning techniques, delivering high performance and robustness in complex scenarios

Nonslop: A Gamified Experiment in Human-AI Collaborative Writing

TutorialsDGX agent

arXiv:2606.12350v1 Announce Type: new Abstract: The rapid proliferation of large language models (LLMs) raises critical questions about human creativity and individual expression in an era of AI-assis

OCSVM-Guided Representation Learning for Unsupervised Anomaly Detection

Model ReleasesDGX agent

arXiv:2507.21164v2 Announce Type: replace-cross Abstract: Unsupervised anomaly detection (UAD) aims to detect anomalies without labeled data, a necessity in many machine learning applications where an

Offline Diffusion Policy for Multi-User Delay-Constrained Scheduling

SafetyDGX agent

arXiv:2501.12942v2 Announce Type: replace Abstract: Effective multi-user delay-constrained scheduling is crucial in various real-world applications, including embodied AI, instant messaging, live stre

OmniBioTwin: A System-of-Twinned-Systems Framework for Health Digital Twins

AgentsDGX agent

arXiv:2606.11264v1 Announce Type: cross Abstract: Health digital twins (HDTs) promise patient-specific modeling and decision support but current approaches remain structurally fragmented: monolithic m

On the Limits of LLM-as-Judge for Scientific Novelty Assessment

Model ReleasesDGX agent

arXiv:2606.12071v1 Announce Type: cross Abstract: LLMs are increasingly used to generate and judge scientific ideas. This makes novelty evaluation a central problem. Full idea evaluation is difficult

On the Optimal Reasoning Length for RL-Trained Language Models

ResearchDGX agent

arXiv:2602.09591v3 Announce Type: replace-cross Abstract: Reinforcement learning substantially improves reasoning in large language models, but it also tends to lengthen chain-of-thought outputs and i

On the Study of Biometric Spoofing Detection using Deep Learning

ResearchDGX agent

arXiv:2606.11505v1 Announce Type: cross Abstract: Biometric systems are increasingly deployed in security applications; however, they remain vulnerable to spoofing attacks, in which attackers exploit

OpenMedReason: Scientific Reasoning Supervision for Medical Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.12169v1 Announce Type: cross Abstract: High-stakes clinical use of large vision-language models (LVLMs) requires reasoning that is grounded in visual evidence and clinical knowledge, not ju

Organize then Retrieve: Hierarchical Memory Navigation for Efficient Agents

AgentsDGX agent

arXiv:2606.11680v1 Announce Type: new Abstract: Large language model (LLM) agents struggle with long-horizon tasks due to their inherent statelessness, requiring all task-relevant information to be en

Ouroboros-Spatial: Closing the Data-Model Loop for Spatial Reasoning

ResearchDGX agent

arXiv:2606.11719v1 Announce Type: cross Abstract: Spatial reasoning remains a persistent challenge for multimodal large language models (MLLMs). Existing approaches largely rely on large-scale, static

Overcoming State Inertia in Full-Duplex Spoken Language Models via Activation Steering

Model ReleasesDGX agent

arXiv:2606.11386v1 Announce Type: cross Abstract: Full-duplex spoken language models (FD-SLMs) enable seamless speech interaction by allowing models to listen and speak simultaneously, yet the interna

Pass@K Policy Optimization: Solving Harder Reinforcement Learning Problems

Model ReleasesDGX agent

arXiv:2505.15201v5 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) algorithms sample multiple n>1 solution attempts for each problem and reward them independently. This optimizes fo

PermDoRA -- Understanding Adapter Interference in Language Models: Limits of Parameter-Space Geometry

Model ReleasesDGX agent

arXiv:2606.11262v1 Announce Type: cross Abstract: Access control in large language models (LLMs) requires modular mechanisms to enable domain-specific behavior without retraining or cross-domain inter

Physics-Distilled Neural Network enabled by Large Language Models for Manufacturing Process-Property Predictive Modeling

ApplicationsDGX agent

arXiv:2606.11605v1 Announce Type: cross Abstract: Predicting process-property relationships in manufacturing is often challenged by high experimental costs and the limited interpretability of complex

Physics-informed generative AI for semiconductor manufacturing: Enforcing hard physical constraints in generative models by construction

AgentsDGX agent

arXiv:2606.11247v1 Announce Type: cross Abstract: Generative models are increasingly used to propose designs, data, and control actions for physical systems, yet many such systems are governed by hard

Planning under Distribution Shifts with Causal POMDPs

TutorialsDGX agent

arXiv:2602.23545v2 Announce Type: replace Abstract: In the real world, planning is often challenged by distribution shifts. As such, a model of the environment obtained under one set of conditions may

PoQ-Judge: A Multi-Architecture Evaluation Framework for Cost-Aware Proof-of-Quality in Decentralized LLM Inference

ResearchDGX agent

arXiv:2606.11196v1 Announce Type: cross Abstract: Decentralized LLM inference networks need lightweight, reference-free quality evaluation for Proof of Quality (PoQ). We present PoQ-Judge, a framework

Position: Hippocampal Explicit Memory Is the Cornerstone for AGI

ResearchDGX agent

arXiv:2606.11245v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities across various tasks, raising expectations for Artificial General Intelligence (A

Position: Stop Anthropomorphizing Intermediate Tokens as Reasoning/Thinking Traces!

TutorialsDGX agent

arXiv:2504.09762v4 Announce Type: replace Abstract: Intermediate token generation (ITG), where a model produces output before the solution, has become a standard method to improve the performance of l

Power Term Polynomial Algebra for Boolean Logic

ResearchDGX agent

arXiv:2603.13854v2 Announce Type: replace-cross Abstract: We introduce power term polynomial algebra, a representation language for Boolean formulae designed to bridge conjunctive normal form (CNF) an

Preregistration for Experiments with AI Agents

AgentsDGX agent

arXiv:2606.11217v1 Announce Type: cross Abstract: The proliferation of large language models (LLMs) and autonomous AI agents has given rise to a rapidly growing methodological paradigm: 'in silico' be

Pretrained self-supervised speech models can recognize unseen consonants

ResearchDGX agent

arXiv:2606.11542v1 Announce Type: cross Abstract: Modern pretrained self-supervised automatic speech recognition models are trained on large-scale audio data to encode speech into contextualized repre

PRInTS: Reward Modeling for Long-Horizon Information Seeking

AgentsDGX agent

arXiv:2511.19314v2 Announce Type: replace Abstract: Information-seeking is a core capability for AI agents, requiring them to gather and reason over tool-generated information across long trajectories

Privacy-Preserving Federated Autoencoder for ECG Anomaly Detection on Edge Devices

ApplicationsDGX agent

arXiv:2606.11556v1 Announce Type: cross Abstract: Continuous electrocardiography (ECG) monitoring could surface rhythm abnormalities before they escalate into cardiovascular events. However, a deploya

ProcessThinker: Enhancing Multi-modal Large Language Models Reasoning via Rollout-based Process Reward

SafetyDGX agent

arXiv:2606.11209v1 Announce Type: cross Abstract: Visual question answering increasingly requires multi-step reasoning. Recent post-training with reinforcement learning under verifiable rewards (RLVR)

ProGRank: Probe-Gradient Reranking to Defend Dense-Retriever RAG from Corpus Poisoning

Model ReleasesDGX agent

arXiv:2603.22934v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) improves large language model applications by grounding generation in retrieved evidence, but also introduces c

PROJECTMEM: A Local-First, Event-Sourced Memory and Judgment Layer for AI Coding Agents

Local AiDGX agent

arXiv:2606.12329v1 Announce Type: new Abstract: AI coding assistants now support a growing share of software work, from quick scripts to production applications. Yet these agents remain largely statel

Quality Adaptive Angular Margin Learning for Respiratory Sound Classification

ResearchDGX agent

arXiv:2606.11915v1 Announce Type: cross Abstract: We present a quality-adaptive angular-margin learning framework that improves feature generalization by enforcing intra-class compactness and inter-cl

Quantifying Subliminal Behavioral Transfer Ratios in Language Model Distillation

Model ReleasesDGX agent

arXiv:2606.11270v1 Announce Type: cross Abstract: Distillation of a language model intended to transfer benign behavior to a student model may also transfer undesirable characteristics, if they are pr

Quantized Stochastic Primal-Dual Methods for Distributed Optimization under Relaxed Global Geometry

ResearchDGX agent

arXiv:2606.11339v1 Announce Type: cross Abstract: We study distributed optimization with stochastic gradients and finite-bit communication modeled by random (unbiased) quantization. We propose q-PDGD,

RAIL: Rethinking Auditory Intelligence in Large Audio-Language Models with a CHC-Grounded Benchmark

Model ReleasesDGX agent

arXiv:2606.11260v1 Announce Type: cross Abstract: Humans process rich auditory environments through tightly integrated cognitive capabilities such as audio perception, audio reasoning, and memory. Des

Reason, Then Re-reason: Cross-view Revisiting Improves Spatial Reasoning

ResearchDGX agent

arXiv:2606.11683v1 Announce Type: cross Abstract: Spatial reasoning from egocentric videos is inherently challenging because the observable evidence is constrained by the camera trajectory. Existing m

Redesign Mixture-of-Experts Routers with Manifold Power Iteration

SafetyDGX agent

arXiv:2606.12397v1 Announce Type: cross Abstract: Router is the cornerstone component to the Mixture-of-Experts models. Serving as expert proxies, the rows of the router matrix compute their similarit

Reinforcement Learning Disrupts Gradient-Based Adversarial Optimization

SafetyDGX agent

arXiv:2606.12251v1 Announce Type: cross Abstract: Gradient-based adversarial attacks remain a dominant threat to deep neural networks (DNNs), as they exploit gradient information to efficiently optimi

RelayFormer: A Unified Local-Global Attention Framework for Scalable Image and Video Manipulation Localization

ResearchDGX agent

arXiv:2508.09459v3 Announce Type: replace-cross Abstract: Visual manipulation localization (VML) aims to identify tampered regions in images and videos, a task that has become increasingly challenging

Reroute, Don't Remove: Recoverable Visual Token Routing for Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.12412v1 Announce Type: cross Abstract: Vision-language models (VLMs) project images into hundreds to thousands of visual tokens, making decoder inference expensive in both attention computa

Resource-Aware LLM Reasoning for Mobile Edge General Intelligence

AgentsDGX agent

arXiv:2509.23248v3 Announce Type: replace Abstract: The rapid advancement of large language models (LLMs) has enabled an emergence of agentic artificial intelligence (AI) with powerful reasoning and a

Risk Under Pressure: Compute-Aware Evaluation of Adversarial Robustness in Language Models

SafetyDGX agent

arXiv:2606.11409v1 Announce Type: cross Abstract: Adversarial robustness evaluations of large language models (LLMs) typically report attack success rate (ASR) under fixed query budgets, implicitly tr

Robust Privacy: Inference-Stage Privacy through Certified Robustness

Model ReleasesDGX agent

arXiv:2601.17360v2 Announce Type: replace-cross Abstract: An adversary observing a model's released prediction can infer sensitive attributes of the queried input, or even reconstruct representatives

RoVE: Rotary Value Embeddings Attention for Relative Position-dependent Value Pathways

Model ReleasesDGX agent

arXiv:2606.11275v1 Announce Type: cross Abstract: Rotary Position Embeddings (RoPE) make attention scores position-relative but leave the value pathway position-blind: the message sent by a value toke

Rule Taxonomy and Evolution in AI IDEs: A Mining and Survey Study

TutorialsDGX agent

arXiv:2606.12231v1 Announce Type: cross Abstract: The adoption of AI-powered Integrated Development Environments (AI IDEs) has introduced 'Rules' as a novel software artifact, allowing developers to p

Runtime Enforcement of Hybrid System Properties

SafetyDGX agent

arXiv:2606.12022v1 Announce Type: cross Abstract: Runtime enforcement has emerged as a promising approach for ensuring the safety of autonomous and cyber-physical systems operating in uncertain and dy

Runtime Skill Audit: Targeted Runtime Probing for Agent Skill Security

AgentsDGX agent

arXiv:2606.11671v1 Announce Type: cross Abstract: Agent skills let LLM agents reuse instructions, resources, tools, and workflows, but they also create a new place for malicious behavior to hide. A sk

Sample-Efficient Hypergradient Estimation for Decentralized Bi-Level Reinforcement Learning

SafetyDGX agent

arXiv:2603.14867v4 Announce Type: replace-cross Abstract: Many strategic decision-making problems, such as environment design for warehouse robots, can be naturally formulated as bi-level reinforcemen

SDQM: Synthetic Data Quality Metric for Object Detection Dataset Evaluation

ResearchDGX agent

arXiv:2510.06596v2 Announce Type: replace-cross Abstract: The performance of machine learning models depends heavily on training data. The scarcity of large-scale, well-annotated datasets poses signif

Search Discipline for Long-Horizon Research Agents

AgentsDGX agent

arXiv:2606.11522v1 Announce Type: new Abstract: Autoresearch agents now propose, evaluate, and select scientific candidates against a metric, and that metric is usually an aggregate reduced over a het

Semantic search for 100M+ galaxy images using AI-generated captions

ResearchDGX agent

arXiv:2512.11982v2 Announce Type: replace-cross Abstract: Finding scientifically interesting phenomena through slow manual labeling campaigns severely limits our ability to explore the billions of gal

Signed Compression Progress on a Sealed Audit is Goodhart-Resistant

SafetyDGX agent

arXiv:2606.11417v1 Announce Type: cross Abstract: Compression progress is a long-standing proposal for intrinsic motivation: reward an agent when its world model becomes better at predicting or compre

SirenFNO: Efficient and Full Frequency Learning of Fourier Neural Operators

Model ReleasesDGX agent

arXiv:2606.11518v1 Announce Type: cross Abstract: Fourier neural operators (FNOs) are effective and efficient surrogates for approximating solutions of PDEs and generalize across discretizations. Howe

Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task

AgentsDGX agent

arXiv:2606.11830v1 Announce Type: new Abstract: Background. Large language models and AI agents are increasingly used to support biomedical research, but native model outputs may omit key analytical s

SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior

AgentsDGX agent

arXiv:2606.11543v1 Announce Type: new Abstract: Agent Skills augment large language model (LLM) agents with procedural knowledge at inference time, but current benchmarks rarely distinguish what a Ski

Small Experiments, Cheaper Decisions: A Case Study in Staged Promotion for Micro-Pretraining

HardwareDGX agent

arXiv:2606.11387v1 Announce Type: cross Abstract: Short pretraining runs can reduce experimental cost, but they can also over-promote configurations that only look strong at tiny budgets. We study an

Soft-Prompt Tuning for Fair and Efficient LLM Benchmark Evaluation

Model ReleasesDGX agent

arXiv:2606.12117v1 Announce Type: cross Abstract: Benchmark scores often misrepresent a large language model's (LLM's) knowledge, because they rely, e.g., on the model's ability to follow specific for

← Previous
1…132133134135136…358
Next →