AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Research

Opt-Verifier: Unleashing the Power of LLMs for Optimization Modeling via Dual-Side Verification

DGX agent

arXiv:2605.29556v1 Announce Type: new Abstract: Building mathematical optimization models is critical in operations research (OR), while it requires substantial human expertise. Recent advancements ha

researcharxiv-cs-ai
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

OptSkills: Learning Generalizable Optimization Skills from Problem Archetypes via Cluster-Based Distillation

DGX agent

arXiv:2605.29829v1 Announce Type: new Abstract: Leveraging Large Language Models (LLMs) to automatically formulate and solve optimization problems from natural language has emerged as an efficient par

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Orthogonal Concept Erasure for Diffusion Models

DGX agent

arXiv:2605.28902v1 Announce Type: new Abstract: Concept erasure has emerged as a promising approach to mitigate undesired or unsafe content in diffusion models, yet existing methods still face signifi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Overcoming Forgetting in LLM Fine-Tuning with Evolution Strategies

DGX agent

arXiv:2605.30148v1 Announce Type: cross Abstract: Evolution Strategies (ES) has recently emerged as a competitive alternative to reinforcement learning (RL) for large language model (LLM) fine-tuning,

model-releasesarxiv-cs-ai
29 May 2026
Applications

P^2RAG: Efficient Privacy-Preserving RAG Service Supporting Arbitrary Top-k Retrieval

DGX agent

arXiv:2603.14778v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) enables large language models to use external knowledge, but outsourcing the RAG service raises privacy c

applicationsarxiv-cs-ai
29 May 2026
Safety

Paper Agents, Paper Gains: An Empirical Analysis of DeFi Investment Agents

DGX agent

arXiv:2605.29174v1 Announce Type: new Abstract: DeFi investment agents, systems that use AI for autonomous on-chain trading, have attained over USD 3 billion in combined token valuations since late 20

safetyarxiv-cs-ai
29 May 2026
Model Releases

Parallax: Parameterized Local Linear Attention for Language Modeling

DGX agent

arXiv:2605.29157v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become the central paradigm in artificial intelligence, yet the core computational primitive of attention has remain

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

ParaTool: Shifting Tool Representations from Context to Parameters

DGX agent

arXiv:2605.29561v1 Announce Type: new Abstract: Tool calling extends large language models (LLMs) by enabling grounded interaction with external executable interfaces, thereby supporting environment-c

model-releasesarxiv-cs-ai
29 May 2026
Applications

PARCEL: Pool-Anchored Resampling with Conditioned Elastic Queries for Efficient Vision-Language Understanding

DGX agent

arXiv:2605.30126v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) map visual inputs into dense token sequences, imposing a quadratic computational bottleneck for inference. Elasti

applicationsarxiv-cs-ai
29 May 2026
Applications

PassNet: Scaling Large Language Models for Graph Compiler Pass Generation

DGX agent

arXiv:2605.29357v1 Announce Type: new Abstract: Modern tensor compilers such as TorchInductor deliver substantial speedups on mainstream models, yet face a systematic performance ceiling on long-tail

applicationsarxiv-cs-ai
29 May 2026
Applications

Persona Conditioning of Brand Recommendations in Retrieval-Augmented Commercial Chat: A Prominence-Stratified Cross-Provider Audit

DGX agent

arXiv:2605.30207v1 Announce Type: new Abstract: The same prompt -- 'best CRM software' -- reaches AI assistants from buyers in widely different contexts: a solo founder, an enterprise VP, a UK SMB own

applicationsarxiv-cs-ai
29 May 2026
Safety

PersonaAgent: Bridging Memory and Action for Personalized LLM Agents

DGX agent

arXiv:2506.06254v2 Announce Type: replace Abstract: Large Language Model (LLM) empowered agents have recently emerged as advanced paradigms that exhibit impressive capabilities in a wide range of doma

safetyarxiv-cs-ai
29 May 2026
Model Releases

Personalized Turn-Level User Conversation Satisfaction Benchmark

DGX agent

arXiv:2605.29711v1 Announce Type: cross Abstract: User satisfaction with AI assistants is highly personalized: the same response may satisfy one user but disappoint another depending on what each user

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

PhoneWorld: Scaling Phone-Use Agent Environments

DGX agent

arXiv:2605.29486v1 Announce Type: cross Abstract: A central bottleneck for phone-use agents is that controllable, reproducible environments covering real mobile behavior are hard to build at scale. Ex

model-releasesarxiv-cs-ai
29 May 2026
Agents

PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions

DGX agent

arXiv:2605.30268v1 Announce Type: cross Abstract: We address the task of generating physically accurate and visually faithful 4D Human-Object Interaction (HOI). Given a static 3D human and target obje

agentsarxiv-cs-ai
29 May 2026
Model Releases

Physics Is All You Need? A Case Study in Physicist-Supervised AI Development of Scientific Software

DGX agent

arXiv:2605.30353v1 Announce Type: new Abstract: Are AI agents tools, co-authors, or researchers? We present a quantified case study (N=1): a physicist supervising an AI coding agent (Claude Code, Sonn

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Planning with the Views via Scene Self-Exploration

DGX agent

arXiv:2605.29563v1 Announce Type: new Abstract: Can VLMs predict how each camera move changes the view, and plan many such moves ahead? We call this capability view planning, requiring (1)understandin

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Pocket-Dentist: On-Device Dental Image Understanding via Efficient Multimodal Large Language Models

DGX agent

arXiv:2605.29299v1 Announce Type: cross Abstract: Evaluations of dental vision-language models remain fragmented across datasets, task definitions and metrics, and often ignore their computational cos

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers

DGX agent

arXiv:2605.30094v1 Announce Type: new Abstract: Poker is a landmark challenge for artificial intelligence. The dominant approach relies on equilibrium solvers built on counterfactual regret minimizati

model-releasesarxiv-cs-ai
29 May 2026
Applications

Position: Text Embeddings Should Capture Implicit Semantics, Not Just Surface Meaning

DGX agent

arXiv:2506.08354v2 Announce Type: replace-cross Abstract: This position paper argues that text embedding research should move beyond surface meaning and embrace implicit semantics as a central modelin

applicationsarxiv-cs-ai
29 May 2026
Safety

Practitioner Beliefs and Behaviors in AI-Enhanced Education: DOT Framework Survey Evidence

DGX agent

arXiv:2605.29041v1 Announce Type: new Abstract: This study reports findings from a cross-sectional survey (n = 72) of higher education practitioners examining beliefs, behaviors, and institutional con

safetyarxiv-cs-ai
29 May 2026
Model Releases

PRAIB: Peer Review AI Benchmark of Behaviour of LLM-Assisted Reviewing

DGX agent

arXiv:2605.29815v1 Announce Type: new Abstract: The growing number of submitted papers has motivated the exploration of Large Language Models (LLMs) as a means to support and augment the peer review p

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Predicting Causal Effects from Natural Language Queries using Structured Representations

DGX agent

arXiv:2605.29631v1 Announce Type: cross Abstract: Randomized controlled trials are a cornerstone of medicine and the social sciences as they enable reliable estimates of causal effects. However, they

model-releasesarxiv-cs-ai
29 May 2026
Tutorials

PrismFlow: Residual Dynamics for Flow Matching in Time-Series Generation

DGX agent

arXiv:2605.28867v1 Announce Type: cross Abstract: Generating high-quality time-series data is challenging because real-world signals often exhibit multimodal patterns and multiscale dynamics, includin

tutorialsarxiv-cs-ai
29 May 2026
Safety

PRO-CUA: Process-Reward Optimization for Computer Use Agents

DGX agent

arXiv:2605.29119v1 Announce Type: new Abstract: Computer use agents (CUAs) have shown strong potential for automating complex digital workflows, yet their training remains constrained by costly live e

safetyarxiv-cs-ai
29 May 2026
Research

Projectional Decoding: Towards Semantic-Aware LLM Generation

DGX agent

arXiv:2605.30054v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate software artifacts across many software engineering (SE) tasks, yet ensuring the semant

researcharxiv-cs-ai
29 May 2026
Model Releases

ProjectionBench: Evaluating Scientific Hypothesis Generation in LLMs Under Progressive Information Disclosure

DGX agent

arXiv:2605.30284v1 Announce Type: new Abstract: Scientific discovery is an inherently creative and uncertain process, requiring reasoning beyond the recall of known knowledge. While many benchmarks ha

model-releasesarxiv-cs-ai
29 May 2026
Safety

Provably Secure Agent Guardrail

DGX agent

arXiv:2605.29251v1 Announce Type: new Abstract: As large language models transition from bounded generative engines to agents with expansive execution privileges, AI going out of control precipitates

safetyarxiv-cs-ai
29 May 2026
Model Releases

PTCG-Bench: Can LLM Agents Master Pokemon Trading Card Game?

DGX agent

arXiv:2605.29653v1 Announce Type: new Abstract: Given a strategically complex board game, human players can quickly learn to devise strategies after playing a few rounds. Autonomous agents require sim

model-releasesarxiv-cs-ai
29 May 2026
Research

Pushing the Limits of Block Rotations in Post-Training Quantization

DGX agent

arXiv:2601.22347v2 Announce Type: replace-cross Abstract: Recent post-training quantization (PTQ) methods have adopted block rotations to diffuse outliers prior to rounding. While this reduces the ove

researcharxiv-cs-ai
29 May 2026
Model Releases

PuzzleClone: A DSL-Powered Framework for Synthesizing Verifiable Data

DGX agent

arXiv:2508.15180v3 Announce Type: replace Abstract: High-quality mathematical and logical datasets with verifiable answers are essential for strengthening the reasoning capabilities of large language

model-releasesarxiv-cs-ai
29 May 2026
Safety

Quantifying and Optimizing Simplicity via Polynomial Representations

DGX agent

arXiv:2605.29823v1 Announce Type: new Abstract: Deep networks often exhibit a preference for 'simple' solutions, and such a simplicity bias is widely believed to play a key role in generalization. Yet

safetyarxiv-cs-ai
29 May 2026
Safety

Quantum-Enhanced Adversarial Robustness in Artificial Intelligence

DGX agent

arXiv:2605.28899v1 Announce Type: cross Abstract: Artificial Intelligence has achieved remarkable success across diverse application domains. However, its vulnerability to adversarial attacks poses si

safetyarxiv-cs-ai
29 May 2026
Safety

Quotient DAGs for Off-Policy Evaluation:Forward-Flow Importance Sampling and Exact Slate Propensities

DGX agent

arXiv:2605.29500v1 Announce Type: cross Abstract: Off-policy evaluation estimates how a target policy would perform using data collected by a different behavior policy, which is crucial when online te

safetyarxiv-cs-ai
29 May 2026
Model Releases

Qwen-VLA: Unifying Vision-Language-Action Modeling across Tasks, Environments, and Robot Embodiments

DGX agent

arXiv:2605.30280v1 Announce Type: cross Abstract: Embodied intelligence is often studied through specialized models for individual tasks such as manipulation or navigation, resulting in fragmented cap

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

RAISE: RAG Design as an Architecture Search Problem

DGX agent

arXiv:2605.30029v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) systems expose numerous design choices spanning query rewriting, chunking, retrieval depth, reranking, and context

model-releasesarxiv-cs-ai
29 May 2026
Agents

Real-rootedness of the Poincare polynomials of overline{mathcal M}_{0,n}: an AI-assisted proof

DGX agent

arXiv:2605.29151v1 Announce Type: cross Abstract: We prove real-rootedness for the Poincare polynomial [ P_n(t)=sum_{i=0}^{n-3} im H^{2i}(overline{mathcal M}_{0,n};Q)t^i ] of the Deligne--Mumford modu

agentsarxiv-cs-ai
29 May 2026
Research

Reasoning about Reasoning: BAPO Bounds on Chain-of-Thought Token Complexity in LLMs

DGX agent

arXiv:2602.02909v2 Announce Type: replace Abstract: Inference-time scaling via chain-of-thought (CoT) reasoning is a major driver of state-of-the-art LLM performance, but it comes with substantial lat

researcharxiv-cs-ai
29 May 2026
Model Releases

Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning

DGX agent

arXiv:2602.00994v2 Announce Type: replace Abstract: Agentic Reinforcement Learning (ARL) trains large language models to interleave reasoning with external tool execution to solve complex tasks. Most

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Reasoning Theater: Disentangling Model Beliefs from Chain-of-Thought

DGX agent

arXiv:2603.05488v4 Announce Type: replace-cross Abstract: We provide evidence of performative chain-of-thought (CoT) in reasoning models, where a model becomes strongly confident in its final answer,

model-releasesarxiv-cs-ai
29 May 2026
Safety

Reasoning While Asking: Transforming Reasoning Large Language Models from Passive Solvers to Proactive Inquirers

DGX agent

arXiv:2601.22139v2 Announce Type: replace-cross Abstract: Reasoning-oriented Large Language Models (LLMs) have achieved remarkable progress with Chain-of-Thought (CoT) prompting, yet they remain funda

safetyarxiv-cs-ai
29 May 2026
Local Ai

Reasoning with Sampling: Cutting at Decision Points

DGX agent

arXiv:2605.30327v1 Announce Type: cross Abstract: Frontier reasoning models are produced by posttraining base language models with reinforcement learning. Recent work has challenged this by showing th

local-aiarxiv-cs-ai
29 May 2026
Safety

ReasonLight: A Multimodal Foundation Model-Enhanced Reinforcement Learning Framework for Zero-Shot Traffic Signal Control

DGX agent

arXiv:2605.29425v1 Announce Type: new Abstract: Reinforcement learning (RL) has shown promise in traffic signal control (TSC). However, its reliance on predefined states limits responsiveness to obser

safetyarxiv-cs-ai
29 May 2026
Model Releases

ReasonOps: Operator Segmentation for LLM Reasoning Traces

DGX agent

arXiv:2605.29192v1 Announce Type: new Abstract: Chain-of-thought traces from large reasoning models can span tens of thousands of tokens, yet we lack a vocabulary for describing their internal structu

model-releasesarxiv-cs-ai
29 May 2026
Safety

Recurrent Structural Policy Gradient for Partially Observable Mean Field Games

DGX agent

arXiv:2602.20141v2 Announce Type: replace Abstract: Mean Field Games (MFGs) provide a principled framework for modelling interactions in large population systems. However, algorithmic progress has bee

safetyarxiv-cs-ai
29 May 2026
Model Releases

Redundant or Necessary? A Benchmark for Detecting Redundant Steps in Agent Trajectories

DGX agent

arXiv:2605.29893v1 Announce Type: new Abstract: LLM-based agents have demonstrated strong capabilities in solving complex tasks through multi-step reasoning and tool use. However, existing evaluation

model-releasesarxiv-cs-ai
29 May 2026
Research

Reinforcement Learning with Robust Rubric Rewards

DGX agent

arXiv:2605.30244v1 Announce Type: cross Abstract: While Reinforcement Learning with Verifiable Rewards (RLVR) is effective for deterministically checkable tasks, many vision-language tasks are partial

researcharxiv-cs-ai
29 May 2026
Research

Rel-MOSS: Towards Imbalanced Relational Deep Learning on Relational Databases

DGX agent

arXiv:2603.07916v2 Announce Type: replace Abstract: In recent advances, to enable a fully data-driven learning paradigm on relational databases (RDB), relational deep learning (RDL) is proposed to str

researcharxiv-cs-ai
29 May 2026
← Previous
1…241242243244245…448
Next →