AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Applications

Enhancing Reinforcement Learning in 3D Environments through Semantic Segmentation: A Case Study in ViZDoom

DGX agent

arXiv:2511.11703v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) in 3D environments with high-dimensional sensory input poses two major challenges: (1) the high memory consumption

applicationsarxiv-cs-ai
29 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Entity-Collision: A Stratified Protocol for Attributing Retrieval Lift in Agent Memory

DGX agent

arXiv:2605.29630v1 Announce Type: cross Abstract: End-to-end agent-memory benchmarks report a single hit@k per retriever, confounding lexical leakage (uncontrolled query/gold/distractor entity overlap

model-releasesarxiv-cs-ai
29 May 2026
Safety

Entropy-KL Divergence-based Token Masking: A Novel Approach for Selective Fine-tuning of Large Language Models

DGX agent

arXiv:2605.29303v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) followed by reinforcement learning (RL) has become a standard post-training paradigm for large language models. This paradi

safetyarxiv-cs-ai
29 May 2026
Safety

EPiC: Efficient Video Camera Control Learning with Precise Anchor-Video Guidance

DGX agent

arXiv:2505.21876v2 Announce Type: replace-cross Abstract: Recent approaches for video generation with camera control often create anchor videos (i.e., rendered videos that approximate desired camera m

safetyarxiv-cs-ai
29 May 2026
Model Releases

ESPO: Early-Stopping Proximal Policy Optimization

DGX agent

arXiv:2605.29860v1 Announce Type: cross Abstract: When a large language model under reinforcement learning commits a wrong reasoning step early in a trajectory, standard algorithms force it to keep ge

model-releasesarxiv-cs-ai
29 May 2026
Agents

Estimating the Empowerment of Language Model Agents

DGX agent

arXiv:2509.22504v3 Announce Type: replace Abstract: As language model (LM) agents become increasingly capable and adopted in real-world applications, there is a growing need for scalable evaluation fr

agentsarxiv-cs-ai
29 May 2026
Research

EvA: An Evidence-First Audio Understanding Paradigm for LALMs

DGX agent

arXiv:2603.27667v2 Announce Type: replace-cross Abstract: Large Audio Language Models (LALMs) still struggle in complex acoustic scenes because they often fail to preserve task-relevant acoustic evide

researcharxiv-cs-ai
29 May 2026
Model Releases

Evaluating Dataset Watermarking for Fine-tuning Traceability of Customized Diffusion Models: A Comprehensive Benchmark and Removal Approach

DGX agent

arXiv:2511.19316v2 Announce Type: replace-cross Abstract: Recent fine-tuning techniques for diffusion models enable them to reproduce specific image sets, such as particular faces or artistic styles,

model-releasesarxiv-cs-ai
29 May 2026
Research

Evaluating Skill and Stability of ArchesWeather and ArchesWeatherGen under Multi-Decadal Climate Simulations

DGX agent

arXiv:2605.29976v1 Announce Type: cross Abstract: We evaluate the climate simulation capabilities of ArchesWeather and ArchesWeatherGen, two machine learning models originally trained for weather fore

researcharxiv-cs-ai
29 May 2026
Research

EviLink: Multi-Path Schema Linking with Uncertainty-Guided Evidence Acquisition for Large-Scale Text-to-SQL

DGX agent

arXiv:2605.29670v1 Announce Type: cross Abstract: Schema linking is a difficult and important step in large-scale Text-to-SQL, where systems must identify a compact yet sufficient schema context from

researcharxiv-cs-ai
29 May 2026
Model Releases

Evolutionary Dynamics of Cooperation in Next-Generation LLM Agent Systems: A Cross-Provider Empirical Extension

DGX agent

arXiv:2605.29874v1 Announce Type: cross Abstract: Do next-generation LLM agents inherit the cooperative biases documented in their predecessors, or does scale and provider diversity reshape equilibriu

model-releasesarxiv-cs-ai
29 May 2026
Safety

Evolutionary Refinement of Generative Graph Topologies: A Hybrid WGAN-GA Approach

DGX agent

arXiv:2605.29161v1 Announce Type: cross Abstract: Generating realistic graph-structured data is challenging due to discrete connectivity, varying graph sizes, and class-specific structural patterns. R

safetyarxiv-cs-ai
29 May 2026
Model Releases

Evolutionary Rule Extraction from Corporate Default Prediction Models

DGX agent

arXiv:2605.29478v1 Announce Type: cross Abstract: Small and medium-sized enterprises (SMEs) represent the majority of firms in most economies and often face financial constraints and higher vulnerabil

model-releasesarxiv-cs-ai
29 May 2026
Agents

Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.29790v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) have emerged as an effective paradigm for complex and long-horizon tasks. However, in real-world tasks, MAS often

agentsarxiv-cs-ai
29 May 2026
Applications

Evolving Features vs Evolving Entire Trees with GP for Interpretable Survival Analysis

DGX agent

arXiv:2605.30119v1 Announce Type: cross Abstract: Survival analysis concerns the task of predicting the time until an event occurs. Often used in the medical field, survival analysis deals with incomp

applicationsarxiv-cs-ai
29 May 2026
Safety

EvoMD-LLM: Learning the Language of Species Evolution in Reactive Molecular Dynamics

DGX agent

arXiv:2605.29394v1 Announce Type: new Abstract: While large language models (LLMs) excel at static scientific reasoning, they struggle to model the temporal structure of dynamic physical processes. We

safetyarxiv-cs-ai
29 May 2026
Research

Extreme dynamic symmetry enables omnidirectional and multifunctional robots

DGX agent

arXiv:2605.29254v1 Announce Type: cross Abstract: Symmetry is a central organizing principle in natural systems, yet its use as a unifying design strategy in robotics has largely remained limited to g

researcharxiv-cs-ai
29 May 2026
Local Ai

FHRFormer: A Self-Supervised Masked Transformer Framework for Fetal Heart Rate Time-Series Inpainting and Forecasting

DGX agent

arXiv:2605.29695v1 Announce Type: new Abstract: Approximately 10% of newborns require assistance to initiate breathing at birth, and around 5% need ventilation support. Fetal heart rate (FHR) monitori

local-aiarxiv-cs-ai
29 May 2026
Local Ai

Finding DoRI: Discovery of Retained Images in Diffusion Models

DGX agent

arXiv:2507.16880v3 Announce Type: replace-cross Abstract: Text-to-image diffusion models (DMs) have achieved remarkable success in image generation. However, concerns about data privacy and intellectu

local-aiarxiv-cs-ai
29 May 2026
Model Releases

FinVerBench: Benchmark Validity and Calibration in Large Language Model Financial Statement Verification

DGX agent

arXiv:2605.29586v1 Announce Type: new Abstract: We introduce FinVerBench, a benchmark and validity study for financial statement verification: determining whether a set of corporate financial statemen

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

First head-to-head comparison of agentic AI applied to the analysis of simulated data of the Einstein Telescope

DGX agent

arXiv:2605.28916v1 Announce Type: cross Abstract: We report a comparison of two state-of-the-art agentic AI systems, Claude Code (Anthropic) and Codex (OpenAI), tasked with autonomously executing a si

model-releasesarxiv-cs-ai
29 May 2026
Applications

Forget Less, Generalize More: Unifying Temporal and Structural Adaptation for Dynamic Graphs

DGX agent

arXiv:2605.29453v1 Announce Type: cross Abstract: Representation learning on dynamic graphs requires capturing complex dependencies that evolve across both time and structure. Existing approaches typi

applicationsarxiv-cs-ai
29 May 2026
Research

FormalEvolve: Neuro-Symbolic Evolutionary Search for Diverse Autoformalization

DGX agent

arXiv:2603.19828v3 Announce Type: replace Abstract: Autoformalization aims to produce formal statements that compile and faithfully preserve the intended meaning of informal mathematics. Yet standard

researcharxiv-cs-ai
29 May 2026
Agents

Formalizing Mathematics at Scale

DGX agent

arXiv:2605.29955v1 Announce Type: new Abstract: We present AutoformBot, a multi-agent system for building an Autoformalized Textbook Library At Scale (Atlas) in Lean 4. AutoformBot orchestrates thousa

agentsarxiv-cs-ai
29 May 2026
Model Releases

FormInv: A Measurement Protocol for Semantic Invariance in Mathematical Reasoning Benchmarks

DGX agent

arXiv:2605.29001v1 Announce Type: cross Abstract: A paraphrase-quality audit of MathCheck (ICLR 2025) detected 4 semantically incorrect paraphrases in 129 groups (3.1%); removing them drops GPT-4o fro

model-releasesarxiv-cs-ai
29 May 2026
Applications

From GPS Points to Travel Patterns: Flexible and Semantic Trajectory Generation with LLMs

DGX agent

arXiv:2605.30014v1 Announce Type: new Abstract: Urban trajectories play a crucial role in modeling urban dynamics and supporting various smart city applications. However, privacy concerns restrict acc

applicationsarxiv-cs-ai
29 May 2026
Research

From Meta-Thought to Execution: Cognitively Aligned Post-Training for Generalizable and Reliable LLM Reasoning

DGX agent

arXiv:2601.21909v2 Announce Type: replace Abstract: Current LLM post-training methods optimize complete reasoning trajectories through Supervised Fine-Tuning (SFT) followed by outcome-based Reinforcem

researcharxiv-cs-ai
29 May 2026
Agents

From Prompts to Context: An Ontology-Driven Framework for Human-Generative AI Collaboration

DGX agent

arXiv:2605.29675v1 Announce Type: cross Abstract: Collaborations with Generative AI often begin with a short prompt and end with an opaque output, leaving implicit who was involved, what task was bein

agentsarxiv-cs-ai
29 May 2026
Research

From Rubrics to Reliable Scores: Evidence-Grounded Text Evaluation with LLM Judges

DGX agent

arXiv:2601.08654v2 Announce Type: replace-cross Abstract: Rubric-based text evaluation increasingly uses large language models (LLMs) as scalable judges, but aligning frozen black-box models with huma

researcharxiv-cs-ai
29 May 2026
Model Releases

From XXLTraffic to EvoXXLTraffic: Scaling Traffic Forecasting to Sensor-Evolving Networks

DGX agent

arXiv:2605.29768v1 Announce Type: new Abstract: Existing traffic forecasting benchmarks assume a fixed sensor set, but real road-sensor networks grow continuously as the road network changes year by y

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Frontier LLM-based agents can overcome the ontology curation bottleneck for natural phenotypes

DGX agent

arXiv:2605.28965v1 Announce Type: new Abstract: Linking free-text phenotype descriptions to ontology terms, typically referred to as phenotype annotation, is essential for the cross-study integration

model-releasesarxiv-cs-ai
29 May 2026
Safety

GDSD: Reinforcement Learning as Guided Denoiser Self-Distillation for Diffusion Language Models

DGX agent

arXiv:2605.29398v1 Announce Type: cross Abstract: Reinforcement learning (RL) can be used to improve the policy (denoiser) of diffusion large language models (dLLMs), while being hindered by the intra

safetyarxiv-cs-ai
29 May 2026
Agents

GenesisFunc: Multi-Agent Data Generation for Accurate and Generalizable Function-Calling

DGX agent

arXiv:2605.28835v1 Announce Type: cross Abstract: Large Language Models (LLMs) extend their capabilities through function-calling (FC), which relies on training data with high quality, diversity, and

agentsarxiv-cs-ai
29 May 2026
Safety

Genetically Aligned Patient Representations Improve Hematological Diagnosis

DGX agent

arXiv:2605.29980v1 Announce Type: cross Abstract: Multimodal alignment of histopathology encoders with transcriptomic and genomic data has been shown to significantly improve performance in downstream

safetyarxiv-cs-ai
29 May 2026
Model Releases

GEO-Bench: Benchmarking Ranking Manipulation in Generative Engine Optimization

DGX agent

arXiv:2605.29107v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly rank products, documents, and recommendations for user queries, which makes manipulating these rankings a gr

model-releasesarxiv-cs-ai
29 May 2026
Research

GiPL: Generative augmented iterative Pseudo-Labeling for Cross-Domain Few-Shot Object Detection

DGX agent

arXiv:2605.29539v1 Announce Type: cross Abstract: Vision-language foundation models have shown promising zero-shot generalization for Cross-Domain Few-Shot Object Detection (CD-FSOD). However, they fa

researcharxiv-cs-ai
29 May 2026
Model Releases

Give it Space! Explicit Disentangling of Positional and Semantic Representations in Encoders

DGX agent

arXiv:2605.30022v1 Announce Type: cross Abstract: Positional encoding (PE) underpins how permutation-invariant Transformers represent sequence order, yet how positional information is processed and st

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Good SFT Optimizes for SFT, Better SFT Prepares for Reinforcement Learning

DGX agent

arXiv:2602.01058v2 Announce Type: replace-cross Abstract: Post-training of reasoning LLMs is a holistic process that typically consists of an offline SFT stage followed by an online reinforcement lear

model-releasesarxiv-cs-ai
29 May 2026
Agents

Governing Technical Debt in Agentic AI Systems

DGX agent

arXiv:2605.29129v1 Announce Type: new Abstract: Agentic AI systems are increasingly being explored as production infrastructure: they reason over multiple steps, call tools, act through workflows, and

agentsarxiv-cs-ai
29 May 2026
Model Releases

GPF-LiveNews: A Streaming Evaluation Protocol for Group-Conditioned Framing in Large Language Models

DGX agent

arXiv:2605.28848v1 Announce Type: cross Abstract: Deployed language models are evaluated in a non-stationary environment: model versions, retrieval layers, safety systems, and real-world inputs all ch

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

GPIC: A Giant Permissive Image Corpus for Visual Generation

DGX agent

arXiv:2605.30341v1 Announce Type: cross Abstract: Studying scalable methods for visual generative modeling requires large, accessible, and stable datasets. We introduce GPIC, a Giant Permissive Image

model-releasesarxiv-cs-ai
29 May 2026
Research

GPS-Enhanced Tourist Mobility Modeling with Seasonal Spatial Priors and LLM-Based Activity Chain Generation

DGX agent

arXiv:2605.29578v1 Announce Type: new Abstract: Tourist mobility poses a distinct challenge for urban transportation planning. Unlike resident commuting, tourist travel is largely non-routine, attract

researcharxiv-cs-ai
29 May 2026
Model Releases

Gram: Assessing sabotage propensities via automated alignment auditing

DGX agent

arXiv:2605.30322v1 Announce Type: cross Abstract: We introduce Gram, an automated alignment auditing framework to assess the propensity of AI agents to engage in sabotage. We evaluate Gemini models ac

model-releasesarxiv-cs-ai
29 May 2026
Safety

Grammar-Aware Literate Generative Mathematical Programming with Compiler-in-the-Loop

DGX agent

arXiv:2601.17670v2 Announce Type: replace-cross Abstract: Mathematical programming is widely employed across various sectors - such as logistics, energy, and workforce planning - to model and solve in

safetyarxiv-cs-ai
29 May 2026
Safety

Graph-Enhanced Policy Optimization in LLM Agent Training

DGX agent

arXiv:2510.26270v2 Announce Type: replace Abstract: Multi-step LLM agents in interactive environments represent a crucial step toward long-horizon decision-making. To train such agents, group-based re

safetyarxiv-cs-ai
29 May 2026
Model Releases

GRASP: Gated Regression-Aware Skill Proposer for Self-Improving LLM Agents

DGX agent

arXiv:2605.29668v1 Announce Type: new Abstract: LLM agents acting in structured environments fail in operational rather than conversational ways, and reliability depends on procedural knowledge of the

model-releasesarxiv-cs-ai
29 May 2026
Safety

GrepSeek: Training Search Agents for Direct Corpus Interaction

DGX agent

arXiv:2605.29307v1 Announce Type: cross Abstract: Large Language Model (LLM) search agents have shown strong promise for knowledge-intensive language tasks through multiple rounds of reasoning and inf

safetyarxiv-cs-ai
29 May 2026
Model Releases

GroundAct: Can LLM Agents Ground Actions in Environmental States?

DGX agent

arXiv:2508.05614v2 Announce Type: replace-cross Abstract: LLM agents achieve 85-96% success on tasks where instructions fully specify the action, but drop to 29-53% when action feasibility depends on

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…242243244245246…452
Next →