AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Safety

Momentum for Reasoning: Dense Intrinsic Signals in Policy Optimization

DGX agent

arXiv:2606.08815v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for eliciting long-chain reasoning in large language models. Ho

safetyarxiv-cs-ai
9 Jun 2026
Research
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

More Bang for the Buck: Improving the Inference of Large Language Models at a Fixed Budget using Reset and Discard (ReD)

DGX agent

arXiv:2601.21522v2 Announce Type: replace-cross Abstract: The performance of large language models (LLMs) on verifiable tasks is usually measured by pass@k, the probability of answering a question cor

researcharxiv-cs-ai
9 Jun 2026
Research

More Yap Less Meaning: Uncovering Self-Improvement Behavior in SLMs

DGX agent

arXiv:2606.08471v1 Announce Type: cross Abstract: Recently, language models have made rapid progress across various domains and applications. However, their capability for self-improvement, i.e., whet

researcharxiv-cs-ai
9 Jun 2026
Research

MOSS-Video-Preview: Toward Real-Time Video Understanding via Cross-Attention

DGX agent

arXiv:2606.07639v1 Announce Type: cross Abstract: Video understanding is shifting from the offline paradigm -- taking a fully recorded video as input and producing a single answer after it ends -- tow

researcharxiv-cs-ai
9 Jun 2026
Research

Multi-planar 2D-U-Net Segmentation of 3D-CT Abdominal Organs augmented by Spatial Occurrence Maps

DGX agent

arXiv:2606.07717v1 Announce Type: cross Abstract: This work proposes a lightweight 2D-U-Net-based framework for segmenting five abdominal organs in large field-of-view 3D CT scans. The method combines

researcharxiv-cs-ai
9 Jun 2026
Agents

Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

DGX agent

arXiv:2606.09748v1 Announce Type: new Abstract: Existing benchmarks for deep research agents (DRAs) assess only single-shot outputs, ignoring a key question: can DRAs improve their reports when guided

agentsarxiv-cs-ai
9 Jun 2026
Safety

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers

DGX agent

arXiv:2601.12263v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) integrate visual and textual knowledge into unified representations that increasingly underpin modern retrieval

safetyarxiv-cs-ai
9 Jun 2026
Applications

Multimodal Group Emotion Recognition In-the-Wild Towards a Privacy-Safe Non-Individual Approach

DGX agent

arXiv:2606.07585v1 Announce Type: cross Abstract: This thesis addresses group emotion recognition (GER) in-the-wild with a focus on privacy preservation. Unlike traditional emotion recognition methods

applicationsarxiv-cs-ai
9 Jun 2026
Model Releases

Multimodal Large Language Models as Synthetic Participants in Video-Based Studies: An Evaluation

DGX agent

arXiv:2606.07541v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have shown strong performance on objective tasks such as video understanding and reasoning. However, it remai

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Muon Learns More Robust and Transferable Features than Adam

DGX agent

arXiv:2606.09658v1 Announce Type: cross Abstract: Muon has recently emerged as a state-of-the-art optimizer for pretraining Large Language Models (LLMs) and vision classifiers. Despite its efficiency

researcharxiv-cs-ai
9 Jun 2026
Safety

Neuro-Symbolic Injection of LTLf Constraints in Autoregressive Reinforcement Learning Policies

DGX agent

arXiv:2606.08312v1 Announce Type: new Abstract: In this work we study offline reinforcement learning (RL) under temporally extended task constraints expressed in Linear Temporal Logic over finite trac

safetyarxiv-cs-ai
9 Jun 2026
Safety

NeuroAlign: Hierarchical Multimodal Fusion of Dynamic and Structural Neuroimaging for MCI Analysis

DGX agent

arXiv:2606.07635v1 Announce Type: cross Abstract: Multimodal neuroimaging fusion of functional MRI (fMRI) and diffusion tensor imaging (DTI) provides complementary information for cognitive impairment

safetyarxiv-cs-ai
9 Jun 2026
Safety

Neutrality Bites: Gender Representation in AI-Generated Animal Stories

DGX agent

arXiv:2606.07969v1 Announce Type: cross Abstract: Gender bias in AI-generated stories is a well-documented problem. While much attention has been paid to reducing or mitigating this bias, it is not al

safetyarxiv-cs-ai
9 Jun 2026
Applications

Next-Token Prediction Learns Generalisable Representations of Sleep Physiology

DGX agent

arXiv:2606.09605v1 Announce Type: new Abstract: Foundation models offer a promising route to compress multi-modal physiological signals into compact representations of human health, with broad applica

applicationsarxiv-cs-ai
9 Jun 2026
Research

No Free Lunch for Synthetic Images under Data Scarcity Conditions

DGX agent

arXiv:2606.07640v1 Announce Type: cross Abstract: This study investigates the trade-offs between fidelity, privacy, and utility in synthetic data generation under conditions of data scarcity and priva

researcharxiv-cs-ai
9 Jun 2026
Safety

NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning

DGX agent

arXiv:2602.21172v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are advancing autonomous driving by replacing modular pipelines with unified end-to-end architectures. However,

safetyarxiv-cs-ai
9 Jun 2026
Tutorials

Not Just After One: Sleep-Inspired Replay Prevents Catastrophic Forgetting After Sequential Tasks

DGX agent

arXiv:2606.08447v1 Announce Type: cross Abstract: One of the critical limitations of artificial neural networks is their lack of ability to continually learn: training on new tasks often leads to inte

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

NutriMLLM: Multimodal Large Language Models for Dietary Micronutrient Analysis

DGX agent

arXiv:2606.08948v1 Announce Type: cross Abstract: Comprehensive estimation of dietary micronutrients from food images could improve clinical nutrition care, but training such models requires large mul

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Observability for Delegated Execution in Agentic AI Systems

DGX agent

arXiv:2606.09692v1 Announce Type: cross Abstract: Delegation-scoped execution is not identifiable from standard observables: audit logs and execution traces can be identical under multiple incompatibl

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Offline Reinforcement Learning for Plasma Control in Nuclear Fusion: Codebase and Benchmark

DGX agent

arXiv:2606.07550v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) offers a promising route for developing plasma controllers from historical tokamak data, since online trial-and-er

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

OmniGameArena: A Unified UE5 Benchmark for VLM Game Agents with Improvement Dynamics

DGX agent

arXiv:2606.09826v1 Announce Type: cross Abstract: Vision-language model (VLM) agents are increasingly deployed in interactive game environments. Yet game benchmarks for VLM agents typically report a s

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

OmniMem: Perturbation-aware Memory Compression for Streaming Audio-Visual LLMs

DGX agent

arXiv:2606.07577v1 Announce Type: new Abstract: Audio-visual large language models (LLMs) hold strong promise for long-form video understanding, yet their long-video inference is fundamentally limited

model-releasesarxiv-cs-ai
9 Jun 2026
Research

On the Complexity of Offline Reinforcement Learning with Q^star-Approximation and Partial Coverage

DGX agent

arXiv:2602.12107v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning under Q^star-approximation and partial coverage, a setting that motivates practical algorithms such as

researcharxiv-cs-ai
9 Jun 2026
Agents

One if by Land, Two if by Sea, Three if by Four Seas, and More to Come -- Values of Perception, Prediction, Communication, and Common Sense in Decision Making

DGX agent

arXiv:2601.06077v2 Announce Type: replace-cross Abstract: This work aims to rigorously define the values of perception, prediction, communication, and common sense in decision making. The defined quan

agentsarxiv-cs-ai
9 Jun 2026
Agents

Online Agent-as-a-Judge: Situation-Generating Evaluation for Interactive Agents

DGX agent

arXiv:2606.08200v1 Announce Type: new Abstract: Evaluating LLM-powered interactive social agents is challenging because socially relevant behaviors depend not only on isolated outputs, but also on pri

agentsarxiv-cs-ai
9 Jun 2026
Research

OnlyDense: Reduced-Order Modeling for Lagrangian simulation

DGX agent

arXiv:2606.09065v1 Announce Type: cross Abstract: In science and engineering, Lagrangian simulation methods such as Smooth Particle Hydrodynamics (SPH) or Material Point Method (MPM) are often employe

researcharxiv-cs-ai
9 Jun 2026
Research

Optical Reasoning: Rethinking Images as an Expressive Reasoning Medium Beyond Text

DGX agent

arXiv:2606.09585v1 Announce Type: new Abstract: Chain-of-Thought (CoT) improves the performance of Large Language Models (LLMs) and has been extended to Multimodal Large Language Models (MLLMs). More

researcharxiv-cs-ai
9 Jun 2026
Research

Optimizing Energy-based Neural Network Training with Coherent Ising Machine

DGX agent

arXiv:2606.09117v1 Announce Type: cross Abstract: While Ising machines serve as advanced physical solvers for the Ising model,enabling applications in combinatorial optimization and neural network tra

researcharxiv-cs-ai
9 Jun 2026
Research

Order Matters: Unveiling the Hidden Impact of Macro Placement Sequences via Proxy-Guided LLM Evolution

DGX agent

arXiv:2606.08904v1 Announce Type: new Abstract: Macro placement is a fundamental step in modern chip physical design, playing a crucial role in determining the solution quality of high-dimensional com

researcharxiv-cs-ai
9 Jun 2026
Local Ai

OSMGraphCLIP: Learning Global Location Representations from OpenStreetMap Graphs

DGX agent

arXiv:2606.08046v1 Announce Type: new Abstract: We present OSMGraphCLIP, a CLIP-style geospatial representation model that learns global location embeddings from freely available OpenStreetMap (OSM) d

local-aiarxiv-cs-ai
9 Jun 2026
Safety

Outage Detection in Self-Healing Smart Grids Using Reinforcement Learning with Spectral Graph Neural Networks

DGX agent

arXiv:2606.07583v1 Announce Type: cross Abstract: Self-healing smart grids can quickly adjust their network configuration during outages to minimize power disruptions. During an outage, several action

safetyarxiv-cs-ai
9 Jun 2026
Safety

Overcoming the Regulatory Bottleneck via Agent-to-Agent Protocols: A Nuclear Case Study

DGX agent

arXiv:2606.07866v1 Announce Type: new Abstract: Regulatory review of advanced nuclear reactor designs routinely spans more than three years and consumes hundreds of millions of dollars in combined reg

safetyarxiv-cs-ai
9 Jun 2026
Safety

Oversight Has a Capacity: Calibrating Agent Guards to a Subjective, Fatiguing Human

DGX agent

arXiv:2606.08919v1 Announce Type: new Abstract: As LLM agents begin to take real, irreversible actions (shell commands, file edits, deploys), the standard safety pattern is a human-in-the-loop approva

safetyarxiv-cs-ai
9 Jun 2026
Agents

PACE: Anytime-Valid Acceptance Tests for Self-Evolving Agents

DGX agent

arXiv:2606.08106v1 Announce Type: new Abstract: Self-evolving agents improve by repeatedly proposing changes to their own prompts, skills, or workflows and keeping those that score higher on a small h

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

PACT: Learning Diverse Diagnostic Strategies via Privileged Synthesis and Branch Consensus

DGX agent

arXiv:2606.08938v1 Announce Type: cross Abstract: Clinical diagnosis requires flexible use of multiple reasoning paradigms under incomplete patient information. Existing LLM-based medical agents show

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

PACT: Self-Evolving Physical Safety Alignment for Diffusion Policies in Embodied Manipulation

DGX agent

arXiv:2606.08414v1 Announce Type: cross Abstract: Diffusion policies have achieved remarkable success in robotic manipulation, yet they often fail to satisfy strict physical constraints required for s

safetyarxiv-cs-ai
9 Jun 2026
Safety

PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR

DGX agent

arXiv:2606.08543v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves large language model reasoning but often suffers from rapid policy-entropy collapse, wher

safetyarxiv-cs-ai
9 Jun 2026
Safety

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

DGX agent

arXiv:2606.07988v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models a

safetyarxiv-cs-ai
9 Jun 2026
Research

Page image classifier fine-tuned on century-spanning archives of scanned documents for further content-specific processing

DGX agent

arXiv:2606.07558v1 Announce Type: cross Abstract: Purpose: Digitization projects in the humanities produce vast, heterogeneous archives of historical documents, making manual sorting impractical at sc

researcharxiv-cs-ai
9 Jun 2026
Local Ai

PAI: Preserving Amplitude Information in Representation-Based Time-Series Anomaly Detection

DGX agent

arXiv:2606.08935v1 Announce Type: cross Abstract: Representation-based time-series anomaly detection algorithms significantly outperform other methods on diverse anomaly detection tasks. However, we n

local-aiarxiv-cs-ai
9 Jun 2026
Safety

PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow

DGX agent

arXiv:2606.07549v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and agent workflows have shown strong promise for computational pathology, yet reliable patc

safetyarxiv-cs-ai
9 Jun 2026
Safety

Payoff scaling shapes cooperation in LLM agents across languages

DGX agent

arXiv:2601.19082v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that negotiate, coordinate, and act on behalf of users. Whether they coo

safetyarxiv-cs-ai
9 Jun 2026
Tutorials

Performative Learning Theory

DGX agent

arXiv:2602.04402v3 Announce Type: replace-cross Abstract: Performative predictions influence the very outcomes they aim to forecast. We study performative predictions that affect a sample (e.g., only

tutorialsarxiv-cs-ai
9 Jun 2026
Model Releases

Personalization Meets Safety:Mechanisms,Risks,and Mitigations in Personalized LLMs

DGX agent

arXiv:2606.09038v1 Announce Type: new Abstract: Large Language Models (LLMs) have enabled increasingly personalized interactions by adapting to users' preferences, contexts, and long-term histories. H

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Phantom transitions in language model fine-tuning

DGX agent

arXiv:2606.07559v1 Announce Type: cross Abstract: Fine-tuning a language model on contexts whose correct completion has a near-synonym competitor often fails silently. The cross-entropy loss decreases

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Pharmacogenomic Knowledge Graph Augmentation for Graph Neural Network-Based Drug-Drug Interaction Prediction

DGX agent

arXiv:2606.07698v1 Announce Type: cross Abstract: Graph neural networks (GNNs) applied to drug-drug interaction (DDI) prediction rely exclusively on molecular structure encoded as SMILES-derived graph

model-releasesarxiv-cs-ai
9 Jun 2026
Research

Physics-Guided Sequence-Based Generative Framework for Acoustic Metamaterial Inverse Design

DGX agent

arXiv:2606.09266v1 Announce Type: cross Abstract: Acoustic metamaterial (AMM) inverse design is particularly challenging for broadband target responses due to acoustic dispersion: a structure that mat

researcharxiv-cs-ai
9 Jun 2026
Research

PhysScene: A Scene Graph Dataset for Scientific Visual Reasoning in Physics Experiments

DGX agent

arXiv:2606.09368v1 Announce Type: cross Abstract: Scene Graphs (SGs) provide structured representations of visual scenes by modeling objects and their pairwise relationships. Despite recent progress,

researcharxiv-cs-ai
9 Jun 2026
← Previous
1…187188189190191…452
Next →