AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Applications

Transplanting, inverting, and preventing a misalignment persona: method-conditional emergent misalignment in Qwen2.5

DGX agent

arXiv:2607.04510v1 Announce Type: cross Abstract: Emergent misalignment (EM) -- the broad misbehaviour a language model acquires after fine-tuning on narrow harmful data -- is mediated in Qwen2.5 mode

applicationsarxiv-cs-ai
7 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TREK: Distill to Explore, Reinforce to Refine

DGX agent

arXiv:2607.05339v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) is effective when the current policy already samples useful reasoning trajectories, but it stalls on hard pr

model-releasesarxiv-cs-ai
7 Jul 2026
Research

TRIAGE: Trustworthy Retrieval Instrumentation And Graph Evaluation

DGX agent

arXiv:2607.03447v1 Announce Type: cross Abstract: Knowledge graphs (KGs) that underpin Graph-based Retrieval-Augmented Generation (Graph-RAG) are increasingly built automatically by LLM-driven extract

researcharxiv-cs-ai
7 Jul 2026
Local Ai

Trust-Region Noise Search for Black-Box Alignment of Diffusion and Flow Models

DGX agent

arXiv:2603.14504v2 Announce Type: replace-cross Abstract: Optimizing the noise samples of diffusion and flow models is an increasingly popular approach to align these models to target rewards at infer

local-aiarxiv-cs-ai
7 Jul 2026
Safety

Trust Region Policy Distillation

DGX agent

arXiv:2607.04751v1 Announce Type: cross Abstract: Big goals are hard to achieve all at once; breaking them into small steps is wiser. We present Trust Region Policy Distillation (TOP-D), which transfo

safetyarxiv-cs-ai
7 Jul 2026
Research

Turbo-Muon: Almost-Orthogonal Pre-Conditioning for Fast Muon Updates

DGX agent

arXiv:2512.04632v2 Announce Type: replace Abstract: Orthogonality-based optimizers, such as Muon, have recently shown strong performance across large-scale training and community-driven efficiency cha

researcharxiv-cs-ai
7 Jul 2026
Safety

Turning Off-Policy Tokens On-Policy: A Plug-in Approach for Improving LLM Alignment

DGX agent

arXiv:2607.04728v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training for large language models (LLMs) follows a efficient paradigm of 'rollout then update', which inevitably res

safetyarxiv-cs-ai
7 Jul 2026
Safety

Two Black Boxes, One Solver: Encoder Probing and Decoder Attribution for Neural Multi-Attribute VRP under Hard-Mask and Recourse Decoders

DGX agent

arXiv:2607.04487v1 Announce Type: cross Abstract: Neural autoregressive solvers for the Multi-Attribute Vehicle Routing Problem (MAVRP) reach competitive cost but offer no per-step justification, a pr

safetyarxiv-cs-ai
7 Jul 2026
Safety

UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning

DGX agent

arXiv:2607.04425v1 Announce Type: cross Abstract: Recent advances in multimodal foundation models and agent systems have driven GUI agents from single-platform task execution toward cross-platform int

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Unbiased Alignment for Large Language Models with Noisy Preferences

DGX agent

arXiv:2607.03248v1 Announce Type: cross Abstract: The alignment of large language models with human preferences is commonly achieved through Reinforcement Learning from Human Feedback or Direct Prefer

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

UNDREAM: Bridging Differentiable Rendering and Photorealistic Simulation for End-to-end Adversarial Attacks

DGX agent

arXiv:2510.16923v3 Announce Type: replace-cross Abstract: Deep learning models deployed in safety critical applications like autonomous driving use simulations to test their robustness against adversa

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Unified Audio Intelligence Without Regressing on Text Intelligence

DGX agent

arXiv:2607.05196v1 Announce Type: cross Abstract: Audio intelligence involves understanding, reasoning about, and generating both audio and speech. In this work, we introduce Nemotron-Labs-Audex-30B-A

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

Unsupervised Behavioral Compression: Learning Low-Dimensional Policy Manifolds through State-Occupancy Matching

DGX agent

arXiv:2603.27044v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is widely recognized as sample-inefficient, a limitation attributable in part to the high dimensionality and

model-releasesarxiv-cs-ai
7 Jul 2026
Research

Unsupervised Features Mining via Activation Geometry

DGX agent

arXiv:2607.04222v1 Announce Type: new Abstract: Interpretability methods aim to reveal the features represented inside large language models (LLMs). Many existing methods begin with labeled examples o

researcharxiv-cs-ai
7 Jul 2026
Applications

Unveiling the Unborn: Advancing Fetal Health Classification through Machine Learning

DGX agent

arXiv:2310.00505v3 Announce Type: replace-cross Abstract: Fetal health classification is a critical task in obstetrics, enabling early identification and management of potential health problems. Howev

applicationsarxiv-cs-ai
7 Jul 2026
Model Releases

URSA: Chemistry-Aware Benchmark for Utilitarian Retrosynthesis Assessment

DGX agent

arXiv:2607.04688v1 Announce Type: cross Abstract: Synthesis planning aiming to find pathways of reactions for a target molecule is one of the most important and challenging tasks in drug discovery. Re

model-releasesarxiv-cs-ai
7 Jul 2026
Research

Using Mechanistic Interpretability to Craft Adversarial Attacks against Large Language Models

DGX agent

arXiv:2503.06269v3 Announce Type: replace-cross Abstract: Traditional white-box methods for creating adversarial perturbations against LLMs typically rely only on gradient computation from the targete

researcharxiv-cs-ai
7 Jul 2026
Tutorials

Verifier-free Test-Time Sampling for Vision-Language-Action Models

DGX agent

arXiv:2510.05681v2 Announce Type: replace-cross Abstract: Vision-Language-Action models (VLAs) have demonstrated remarkable performance in robot control. However, they remain fundamentally limited in

tutorialsarxiv-cs-ai
7 Jul 2026
Model Releases

VERITAS: Towards a General-Purpose Replication Tool for Scientific Research

DGX agent

arXiv:2607.02931v1 Announce Type: new Abstract: AI tools are accelerating scientific publication while the systems that review it struggle to keep up, and independent verification of published researc

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

VideoSearcher: Empowering Video Deep Research with Multi-Tool Agentic Reasoning via Reinforcement Learning

DGX agent

arXiv:2607.02927v1 Announce Type: cross Abstract: Video understanding is moving beyond closed-context perception toward open-world evidence exploration, a paradigm formalized as Video Deep Research (V

model-releasesarxiv-cs-ai
7 Jul 2026
Research

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

DGX agent

arXiv:2607.03657v1 Announce Type: cross Abstract: Gloss-free Sign Language Translation (SLT) translates sign language videos into spoken-language sentences without gloss annotations, avoiding costly l

researcharxiv-cs-ai
7 Jul 2026
Research

Vision Token Manipulation Attacks on Cloud-Edge Inference of Large Vision-Language Models

DGX agent

arXiv:2607.02819v1 Announce Type: cross Abstract: Cloud-edge Large Vision-Language Model (LVLM) inference enables efficient deployment by splitting computation between edge devices and cloud servers.

researcharxiv-cs-ai
7 Jul 2026
Safety

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models

DGX agent

arXiv:2509.25533v2 Announce Type: replace-cross Abstract: As Vision Language Models (VLMs) are deployed across safety-critical applications, understanding and controlling their behavioral patterns has

safetyarxiv-cs-ai
7 Jul 2026
Safety

VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models

DGX agent

arXiv:2508.08521v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are increasingly being used in a broad range of applications, bringing their security and behavioral control to

safetyarxiv-cs-ai
7 Jul 2026
Research

VISTA: Auditing Semantic Divergence in Vision-Language Models

DGX agent

arXiv:2607.02995v1 Announce Type: cross Abstract: Vision-language models can exhibit visual concept-conditioned divergence: given images containing demographic features, corporate logos, or ideologica

researcharxiv-cs-ai
7 Jul 2026
Safety

VLA Grounder: Language-Conditioning Space Optimization for Black-Box VLA Models

DGX agent

arXiv:2607.04517v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are commonly treated as end-to-end action policies conditioned on natural-language task descriptions. In practice, h

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Walrus: A Cross-Domain Foundation Model for Continuum Dynamics

DGX agent

arXiv:2511.15684v2 Announce Type: replace-cross Abstract: Foundation models have transformed machine learning for language and vision, but achieving comparable impact in physical simulation remains a

model-releasesarxiv-cs-ai
7 Jul 2026
Hardware

Wan-Streamer v0.2: Higher Resolution, Same Latency

DGX agent

arXiv:2607.04443v1 Announce Type: cross Abstract: We present Wan-Streamer v0.2, a latency-preserving upgrade of the native-streaming, end-to-end audio-visual interaction model. v0.2 keeps the v0.1 mod

hardwarearxiv-cs-ai
7 Jul 2026
Research

Wasserstein Residuals: Learning Gradient Flows from Population Dynamics

DGX agent

arXiv:2607.04738v1 Announce Type: cross Abstract: Reconstructing population dynamics is a central problem in the physical and data sciences. Often, the dynamics are modeled as a Wasserstein gradient f

researcharxiv-cs-ai
7 Jul 2026
Research

Wavelet Scattering Transform for Interpretable Schizophrenia Biomarker Discovery and Classification from Resting-State EEG

DGX agent

arXiv:2607.05282v1 Announce Type: cross Abstract: Schizophrenia is a debilitating neuropsychiatric disorder characterized by profound cortical network dysregulation, for which objective, clinically tr

researcharxiv-cs-ai
7 Jul 2026
Safety

Weak-to-Strong Generalization via Direct On-Policy Distillation

DGX agent

arXiv:2607.05394v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a powerful recipe for improving language-model reasoning, but it is expensive to repeat on ev

safetyarxiv-cs-ai
7 Jul 2026
Agents

Web-CogReasoner: Towards Multimodal Knowledge-Induced Cognitive Reasoning for Web Agents

DGX agent

arXiv:2508.01858v3 Announce Type: replace-cross Abstract: Multimodal large-scale models have significantly advanced the development of web agents, enabling perception and interaction with digital envi

agentsarxiv-cs-ai
7 Jul 2026
Tutorials

What Does a Discrete Diffusion Model Learn?

DGX agent

arXiv:2607.05381v1 Announce Type: cross Abstract: What does a discrete diffusion model learn: a denoiser, a score ratio, or a bridge plug-in predictor? At the level of jump rates, these are one object

tutorialsarxiv-cs-ai
7 Jul 2026
Applications

What is Left for Us? Second Scholarship Against the Degradation of Research by AI

DGX agent

arXiv:2607.04049v1 Announce Type: new Abstract: We argue that generative AI can degrade research by eroding the very practices through which scholarly judgement is formed and academic trust is built.

applicationsarxiv-cs-ai
7 Jul 2026
Model Releases

When Aggregate Alignment Misleads: Auditing Policy Repair Without Per-State Expert Actions

DGX agent

arXiv:2607.03386v1 Announce Type: new Abstract: Agentic AI systems are increasingly used to edit, refine, and repair decision policies, but evaluating these edits is difficult when per-state expert ac

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Claws Remember but Do Not Tell: Stealthy Memory Injection in Persistent Personal Agents

DGX agent

arXiv:2607.05189v1 Announce Type: cross Abstract: Persistent personal agents combine long-term memory with access to users' external environments, enabling personalized foreground assistance and proac

model-releasesarxiv-cs-ai
7 Jul 2026
Research

When Does Small Data Work? Accuracy and Efficiency Trade-offs Between Tabular Foundation Models and Conventional Methods for Crowd-State Classification at Hajj and Umrah

DGX agent

arXiv:2607.04013v1 Announce Type: cross Abstract: Learning from few labeled examples is a central challenge in tabular machine learning, and it becomes the binding constraint in domains where labeling

researcharxiv-cs-ai
7 Jul 2026
Model Releases

When is a System Discoverable from Data? Discovery Requires Chaos

DGX agent

arXiv:2511.08860v2 Announce Type: replace-cross Abstract: The deep learning revolution has spurred a rise in advances of using AI in sciences. Within physical sciences the main focus has been on disco

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Rubrics Fail: Error Enumeration as Reward in Reference-Free RL Post-Training for Virtual Try-On

DGX agent

arXiv:2603.05659v3 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) and Rubrics as Rewards (RaR) have driven strong gains in domains with clear correctness

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

When Simpler Is Better: Evaluating Translation Pipelines for Medieval Latin Manuscripts

DGX agent

arXiv:2607.03836v1 Announce Type: cross Abstract: Despite remarkable progress in machine translation, Vision Language Models (VLMs) struggle on historical manuscripts, a domain that stresses core Natu

model-releasesarxiv-cs-ai
7 Jul 2026
Research

Where do LLMs Fall Short in CBT-Guided Affective Reasoning?

DGX agent

arXiv:2607.02885v1 Announce Type: cross Abstract: Cognitive Behavioral Therapy (CBT) provides a structured framework for understanding a user's mental state by examining the interaction between cognit

researcharxiv-cs-ai
7 Jul 2026
Model Releases

Which Algorithm Specification Formats Help Language Models Implement Machine Learning Algorithms?

DGX agent

arXiv:2607.03158v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to implement algorithms from research manuscripts, but papers often leave implementation choices im

model-releasesarxiv-cs-ai
7 Jul 2026
Research

Why Pure Reasoning is Not Enough: Nature as the Source of Mathematical Innovation

DGX agent

arXiv:2607.04505v1 Announce Type: new Abstract: We advance the hypothesis that human mathematical reasoning, constrained by both the undecidability and the computational intractability of even modest

researcharxiv-cs-ai
7 Jul 2026
Research

Why3-py: A Tool for Formal Verification of Hypothesis Testing and Meta-Analysis in Python

DGX agent

arXiv:2607.03951v1 Announce Type: cross Abstract: The reproducibility crisis in scientific research has received widespread recognition, thereby increasing the importance of meta-analyses that integra

researcharxiv-cs-ai
7 Jul 2026
Research

Worldscape-MoE: A Unified Mixture-of-Experts World Model for Scalable Heterogeneous Action Control

DGX agent

arXiv:2607.03964v1 Announce Type: cross Abstract: World models are rapidly becoming a core infrastructure for embodied intelligence and interactive agents: they provide controllable simulators in whic

researcharxiv-cs-ai
7 Jul 2026
Agents

Your Agent's Memories Are Not Its Own: Forged Reasoning Attacks on LLM Agent Memory and Defenses

DGX agent

arXiv:2607.05029v1 Announce Type: cross Abstract: Persistent memory has enabled large language model (LLM) agents to store factual knowledge, prior decisions, reasoning histories, tool usage informati

agentsarxiv-cs-ai
7 Jul 2026
Agents

A Dual-Helix Governance Approach Towards Reliable Agentic Artificial Intelligence for WebGIS Development

DGX agent

arXiv:2603.04390v2 Announce Type: replace Abstract: WebGIS development requires consistency, yet agentic AI often fails due to LLM context constraints, forgetting, stochasticity, instruction failure,

agentsarxiv-cs-ai
3 Jul 2026
Local Ai

A General Neural Backbone for Mixed-Integer Linear Optimization via Dual Attention

DGX agent

arXiv:2601.04509v2 Announce Type: replace Abstract: Mixed-integer linear programming (MILP) is a foundational framework for combinatorial optimization across science and engineering, but remains hard

local-aiarxiv-cs-ai
3 Jul 2026
← Previous
1…116117118119120…448
Next →