AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Safety

On-the-fly Repulsion in the Contextual Space for Rich Diversity in Diffusion Transformers

DGX agent

arXiv:2603.28762v2 Announce Type: replace-cross Abstract: Modern Text-to-Image (T2I) diffusion models have achieved remarkable semantic alignment, yet they often suffer from a significant lack of vari

safetyarxiv-cs-ai
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval

DGX agent

arXiv:2606.04391v1 Announce Type: new Abstract: Language agents increasingly rely on reusable skills to improve multi-step web automation across related tasks. A growing line of work studies online sk

model-releasesarxiv-cs-ai
4 Jun 2026
Research

OpenRFM: Dissecting Relational In-Context Learning

DGX agent

arXiv:2606.04320v1 Announce Type: cross Abstract: Relational Foundation Models (RFMs) promise a single pre-trained predictor that, given any relational database, returns predictions in one forward pas

researcharxiv-cs-ai
4 Jun 2026
Model Releases

Optical-Guided Neural Collapse for SAR Few-Shot Class Incremental Learning

DGX agent

arXiv:2606.04528v1 Announce Type: cross Abstract: Few-shot class-incremental learning (FSCIL) in synthetic aperture radar imagery presents unique challenges due to severe data scarcity and SAR-specifi

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

Outcome-Based RL Provably Leads Transformers to Reason, but Only With the Right Data

DGX agent

arXiv:2601.15158v4 Announce Type: replace-cross Abstract: Transformers trained via Reinforcement Learning (RL) with outcome-based supervision can spontaneously develop the ability to generate intermed

safetyarxiv-cs-ai
4 Jun 2026
Research

Overview of the EReL@MIR 2025 Multimodal Document Retrieval Challenge (Track 1)

DGX agent

arXiv:2606.04240v1 Announce Type: cross Abstract: Retrieval over visually-rich documents, pages that interleave text with figures, tables, and charts, is essential for multimodal retrieval-augmented g

researcharxiv-cs-ai
4 Jun 2026
Tutorials

ParetoPilot: Zero-Surrogate Offline Multi-Objective Optimization via Infer-Perturb-Guide Diffusion

DGX agent

arXiv:2606.04468v1 Announce Type: cross Abstract: Offline multi-objective optimization (Offline MOO) aims to discover novel Pareto-optimal designs based on static datasets without expensive environmen

tutorialsarxiv-cs-ai
4 Jun 2026
Agents

Parthenon Law: A Self-Evolving Legal-Agent Framework

DGX agent

arXiv:2606.04602v1 Announce Type: new Abstract: As agents grow more capable, legal-domain LLM agents promise to turn document-heavy matters into reviewable work products -- yet reliable deployment fac

agentsarxiv-cs-ai
4 Jun 2026
Safety

PerceptTwin: Semantic Scene Reconstruction for Iterative LLM Planning and Verification

DGX agent

arXiv:2606.04226v1 Announce Type: cross Abstract: Simulation environments are useful for both robot policy learning and planning verification and validation. Traditionally, the process of creating a s

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

PersistBench: When Should Long-Term Memories Be Forgotten by LLMs?

DGX agent

arXiv:2602.01146v2 Announce Type: replace Abstract: Conversational assistants are increasingly integrating long-term memory with large language models (LLMs). This persistence of memories, e.g., the u

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

Physics-Informed Machine Learning for Short-Term Flood Prediction

DGX agent

arXiv:2606.04143v1 Announce Type: cross Abstract: Accurate flood forecasting is essential for mitigating disaster risks and protecting communities. However, purely data-driven machine learning models

safetyarxiv-cs-ai
4 Jun 2026
Research

Physics-Informed Neural Engine Sound Modeling with Differentiable Pulse-Train Synthesis

DGX agent

arXiv:2603.09391v2 Announce Type: replace-cross Abstract: Engine sounds originate from sequential exhaust pressure pulses rather than sustained harmonic oscillations. While neural synthesis methods ty

researcharxiv-cs-ai
4 Jun 2026
Safety

Plan First, Judge Later, Run Better: A DMAIC-Inspired Agentic System for Industrial Anomaly Detection

DGX agent

arXiv:2606.04599v1 Announce Type: new Abstract: Large language model (LLM) agents have shown promise in automating complex data-analysis workflows, but their reliable deployment remains challenging in

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Plan, Watch, Recover: A Benchmark and Architectures for Proactive Procedural Assistance

DGX agent

arXiv:2606.04970v1 Announce Type: cross Abstract: We envision a proactive multi-modal assistant system which gives users real-time step-by-step guidance on a procedural task, autonomously deciding ext

model-releasesarxiv-cs-ai
4 Jun 2026
Research

Platonic Transformers: A Solid Choice For Equivariance

DGX agent

arXiv:2510.03511v3 Announce Type: replace-cross Abstract: While widespread, Transformers lack inductive biases for geometric symmetries common in science and computer vision. Existing equivariant meth

researcharxiv-cs-ai
4 Jun 2026
Safety

POLARIS: Guiding Small Models to Write Long Stories

DGX agent

arXiv:2606.04095v1 Announce Type: cross Abstract: Small open-weight models struggle at long-form creative writing: their generated stories either fall far short of the requested length, or their quali

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

PoliticsBench: Benchmarking Political Values in Large Language Models with Multi-Turn Roleplay

DGX agent

arXiv:2603.23841v2 Announce Type: replace-cross Abstract: While Large Language Models (LLMs) are increasingly used as primary sources of information, their potential for political bias may impact thei

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Position: Deployed Reinforcement Learning should be Continual

DGX agent

arXiv:2606.04029v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has received increasing attention and adoption in real-world use cases. Most of these systems follow a train-then-fix para

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

Proof-Carrying Agent Actions: Model-Agnostic Runtime Governance for Heterogeneous Agent Systems

DGX agent

arXiv:2606.04104v1 Announce Type: cross Abstract: Agent systems execute through runtimes with very different control points: local coding tools, framework SDKs, managed agent platforms, API gateways,

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Provably Auditable and Safe LLM Agents from Human-Authored Ontologies

DGX agent

arXiv:2606.04903v1 Announce Type: cross Abstract: We introduce the LLM agent architecture Agentic Redux, intended for use with nontrivial problem domains that require linear auditability. Using the ty

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

QO-Bench: Diagnosing Query-Operator-Preserving Retrieval over Typed Event Tuples

DGX agent

arXiv:2606.04646v1 Announce Type: cross Abstract: Many real-world questions over business, legal, and scientific corpora are natural-language versions of database-style queries over records latent in

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Quantum entanglement provides a competitive advantage in adversarial games

DGX agent

arXiv:2603.10289v2 Announce Type: replace-cross Abstract: Whether uniquely quantum resources confer advantages in fully classical, competitive environments remains an open question. Competitive zero-s

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

QuBLAST: A Framework for Quantizing Large Language Models with Block-Level Compression Approach and Activation Scaling Strategy

DGX agent

arXiv:2606.04620v1 Announce Type: cross Abstract: LLMs have become the state-of-the-art algorithms for solving NLP tasks. However, they typically come at huge computational and memory costs, thus maki

model-releasesarxiv-cs-ai
4 Jun 2026
Local Ai

R-APS: Compositional Reasoning and In-Context Meta-Learning for Constrained Design via Reflective Adversarial Pareto Search

DGX agent

arXiv:2606.04823v1 Announce Type: new Abstract: Large language models (LLMs) are fluent on open-ended tasks, yet in agentic settings, where a system must plan, use tools, and act over extended horizon

local-aiarxiv-cs-ai
4 Jun 2026
Research

R3G: A Reasoning-Retrieval-Reranking Framework for Vision-Centric Answer Generation

DGX agent

arXiv:2602.00104v3 Announce Type: replace-cross Abstract: Vision-centric retrieval for VQA requires retrieving images to supply missing visual cues and integrating them into the reasoning process. How

researcharxiv-cs-ai
4 Jun 2026
Research

Real-Time Automatic License Plate Recognition Using YOLOv8, SORT Tracking, and Temporal Data Interpolation

DGX agent

arXiv:2606.04684v1 Announce Type: cross Abstract: The real-time hardships of video processing seriously limit the usage of Automatic License Plate Recognition (ALPR) with application in dynamic traffi

researcharxiv-cs-ai
4 Jun 2026
Local Ai

Reasoning or Fluency? Dissecting Probabilistic Confidence in Best-of-N Selection

DGX agent

arXiv:2601.13735v2 Announce Type: replace Abstract: Probabilistic confidence metrics are increasingly adopted as proxies for reasoning quality in Best-of-N selection, under the assumption that higher

local-aiarxiv-cs-ai
4 Jun 2026
Local Ai

Recover-LoRA for Aggressive Quantization: Reclaiming Accuracy in 2-Bit Language Models via Low-Rank Adaptation with Knowledge Distillation on Synthetic Data

DGX agent

arXiv:2606.04238v1 Announce Type: cross Abstract: Aggressive weight quantization to 2-bit precision offers substantial throughput and memory gains for large language model (LLM) inference, but typical

local-aiarxiv-cs-ai
4 Jun 2026
Safety

Reinforcement Learning from Rich Feedback with Distributional DAgger

DGX agent

arXiv:2606.05152v1 Announce Type: cross Abstract: Reasoning models have advanced rapidly, but the dominant reinforcement learning from verifiable rewards (RLVR) recipe remains surprisingly narrow: sam

safetyarxiv-cs-ai
4 Jun 2026
Safety

Reproducing, Analyzing, and Detecting Reward Hacking in Rubric-Based Reinforcement Learning

DGX agent

arXiv:2606.04923v1 Announce Type: cross Abstract: Rubric-based reinforcement learning (RL) uses an LLM-as-a-Judge (LaaJ) to score model outputs according to rubrics as rewards. However, policy models

safetyarxiv-cs-ai
4 Jun 2026
Safety

Rethinking Sales Lead Scoring with LLM-based Hierarchical Preference Ranking

DGX agent

arXiv:2606.04387v1 Announce Type: cross Abstract: Sales lead conversion in high-stakes domains (e.g., automotive, real estate) differs fundamentally from e-commerce recommendation due to prolonged dec

safetyarxiv-cs-ai
4 Jun 2026
Research

Revisiting Model Stitching In the Foundation Model Era

DGX agent

arXiv:2603.12433v3 Announce Type: replace-cross Abstract: Model stitching, connecting early layers of one model (source) to later layers of another (target) via a light stitch layer, has served as a p

researcharxiv-cs-ai
4 Jun 2026
Model Releases

Revisiting Vul-RAG: Reproducibility and Replicability of RAG-based Vulnerability Detection with Open-Weight Models

DGX agent

arXiv:2606.04739v1 Announce Type: cross Abstract: Large language models (LLMs) have shown strong potential for automated software vulnerability detection, particularly in retrieval-augmented generatio

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Rollout-Level Advantage-Prioritized Experience Replay for GRPO

DGX agent

arXiv:2606.04560v1 Announce Type: cross Abstract: Reinforcement learning from verifiable rewards with GRPO is a standard approach for post-training reasoning LLMs. It remains sample inefficient. Each

model-releasesarxiv-cs-ai
4 Jun 2026
Local Ai

RowNet: A Memory Transformer for Tabular Regression

DGX agent

arXiv:2606.04445v1 Announce Type: cross Abstract: Real estate valuation is a structured regression problem in which prices are governed by heterogeneous feature types, sparse regional effects, nonline

local-aiarxiv-cs-ai
4 Jun 2026
Safety

RUBAS: Rubric-Based Reinforcement Learning for Agent Safety

DGX agent

arXiv:2606.04051v1 Announce Type: cross Abstract: The evolution of LLMs into tool-enabled agents creates a new class of safety challenges associated with real-world execution rather than simple text g

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety

DGX agent

arXiv:2603.10044v2 Announce Type: replace-cross Abstract: A safety score earned on a benchmark need not predict how the same model behaves once it is wrapped in an agentic scaffold the benchmark never

model-releasesarxiv-cs-ai
4 Jun 2026
Research

SaliMory: Orchestrating Cognitive Memory for Conversational Agents

DGX agent

arXiv:2606.04120v1 Announce Type: cross Abstract: Conversational agents that serve as lifelong companions must maintain persistent memory across all interactions. However, simply expanding context win

researcharxiv-cs-ai
4 Jun 2026
Model Releases

SAM 3D: 3Dfy Anything in Images

DGX agent

arXiv:2511.16624v2 Announce Type: replace-cross Abstract: We present SAM 3D, a generative model for visually grounded 3D object reconstruction, predicting geometry, texture, and layout from a single i

model-releasesarxiv-cs-ai
4 Jun 2026
Hardware

Scaling Novel Graph Generation via Lightweight Structure-Guided Autoregressive Models

DGX agent

arXiv:2606.04287v1 Announce Type: cross Abstract: Generating realistic and diverse graphs is a key problem in machine learning, with applications in molecular discovery, circuit design, cybersecurity,

hardwarearxiv-cs-ai
4 Jun 2026
Safety

Scaling Self-Evolving Agents via Parametric Memory

DGX agent

arXiv:2606.04536v1 Announce Type: new Abstract: Existing memory-augmented LLM agents store past experience exclusively in prompt space, as textual summaries or retrieved passages, while keeping model

safetyarxiv-cs-ai
4 Jun 2026
Safety

Scenario Generation for Risk-Aware Reinforcement Learning with Probably Approximately Safe Guarantees

DGX agent

arXiv:2606.04812v1 Announce Type: cross Abstract: Guaranteeing safety is critical to the deployment of reinforcement learning (RL) agents in the real-world, especially as policies learned using deep R

safetyarxiv-cs-ai
4 Jun 2026
Research

SCI-PRM: A Tool Aware Process Reward Model for Scientific Reasoning Verification

DGX agent

arXiv:2606.04579v1 Announce Type: new Abstract: While Process Reward Models (PRMs) have achieved remarkable success in mathematical reasoning, their application in complex scientific domains-such as b

researcharxiv-cs-ai
4 Jun 2026
Safety

Selective Coupling of Decoupled Informative Regions: Masked Attention Alignment for Data-Free Quantization of Vision Transformers

DGX agent

arXiv:2606.04373v1 Announce Type: cross Abstract: Data-Free Quantization (DFQ) addresses data security concerns by synthesizing samples, without accessing real data. It has garnered increasing attenti

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Self-Evolving Deep Research via Joint Generation and Evaluation

DGX agent

arXiv:2606.04507v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become increasingly adopted in daily applications, with deep research standing out as a particularly important capab

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Self-Reflective APIs: Structure Beats Verbosity for AI Agent Recovery

DGX agent

arXiv:2606.05037v1 Announce Type: cross Abstract: When an AI agent calls an API and hits a validation error, it needs more than what went wrong -- it needs what to do next. A self-reflective API retur

agentsarxiv-cs-ai
4 Jun 2026
Agents

Semantic Constraint Synthesis for Adaptive Trajectory Optimization via Large Language Models

DGX agent

arXiv:2606.04123v1 Announce Type: cross Abstract: Trajectory optimization is a critical component for enabling safe and reliable autonomous operations in space exploration. As space missions increase

agentsarxiv-cs-ai
4 Jun 2026
Safety

Semiparametric Preference Optimization: Your Language Model is Secretly a Single-Index Model

DGX agent

arXiv:2512.21917v3 Announce Type: replace-cross Abstract: Policy alignment to preference data typically assumes a known link function between observed preferences and latent rewards (e.g., Bradley-Ter

safetyarxiv-cs-ai
4 Jun 2026
← Previous
1…201202203204205…448
Next →