AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Research

Re-Centering Humans in LLM Personalization

DGX agent

arXiv:2606.06614v1 Announce Type: cross Abstract: Despite growing interest, most evaluations of large language models' (LLMs') personalization abilities have relied on synthetic data. It remains uncle

researcharxiv-cs-ai
8 Jun 2026
Safety

Re-imagining ISO 26262 in the Age of Autonomous Vehicles: Enhancing Controllability through Transferability and Predictability

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.07437v1 Announce Type: cross Abstract: The ISO 26262 standard defines functional safety for road vehicles through risk assessments based on Severity, Exposure, and Controllability, grounded

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

ReclAIm: A Multi-Agent Framework for Monitoring and Correcting Performance Decline in Medical Imaging AI

DGX agent

arXiv:2510.17004v2 Announce Type: replace-cross Abstract: Purpose: To develop and evaluate a multi-agent framework (ReclAIm) for automated monitoring, detection, and correction of performance decline

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

REMEDI: A Benchmark for Retention and Unlearning Evaluation in Multi-label Clinical Disease Inference

DGX agent

arXiv:2606.07141v1 Announce Type: cross Abstract: Language models trained for clinical disease inference are trained on patient data, which may include sensitive and private information, and data owne

model-releasesarxiv-cs-ai
8 Jun 2026
Research

RePo: Language Models with Context Re-Positioning

DGX agent

arXiv:2512.14391v3 Announce Type: replace-cross Abstract: In-context learning is fundamental to modern Large Language Models (LLMs); however, prevailing architectures impose a rigid and fixed contextu

researcharxiv-cs-ai
8 Jun 2026
Research

Rethinking Genomic Modeling Through Optical Character Recognition

DGX agent

arXiv:2602.02014v2 Announce Type: replace-cross Abstract: Recent genomic foundation models largely adopt large language model architectures that treat DNA as a one-dimensional token sequence. However,

researcharxiv-cs-ai
8 Jun 2026
Model Releases

RETROSPECT: RETROsynthesis via Sequential Prediction, and Chemically Transformed-ranking

DGX agent

arXiv:2606.07181v1 Announce Type: cross Abstract: Single-step retrosynthesis needs both accurate first-ranked suggestions and candidate lists that are rich enough for downstream selection. We study th

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

Robust Driving Control for Autonomous Vehicles: An Intelligent General-sum Constrained Adversarial Reinforcement Learning Approach

DGX agent

arXiv:2510.09041v3 Announce Type: replace-cross Abstract: Deep reinforcement learning (DRL) has demonstrated remarkable success in developing autonomous driving policies. However, its vulnerability to

safetyarxiv-cs-ai
8 Jun 2026
Safety

SafeGene: Reusable Adapters for Transferable Safety Alignment

DGX agent

arXiv:2606.06519v1 Announce Type: new Abstract: Open-weight LLMs are increasingly fine-tuned into customized assistants, but downstream fine-tuning can weaken safety alignment and make models more vul

safetyarxiv-cs-ai
8 Jun 2026
Agents

SCALE: Scalable Cross-Attention Learning with Extrapolation for Agentic Workflow Scheduling

DGX agent

arXiv:2606.06820v1 Announce Type: cross Abstract: Agentic Large Language Model (LLM) systems decompose complex tasks into workflow Directed Acyclic Graphs (DAGs) whose primitives must be scheduled on

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

ScenicRules: An Autonomous Driving Benchmark with Multi-Objective Specifications and Abstract Scenarios

DGX agent

arXiv:2602.16073v2 Announce Type: replace-cross Abstract: Developing autonomous driving systems for complex traffic environments requires balancing multiple objectives, such as avoiding collisions, ob

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

SCOUT: Semantic scene COverage via Uncertainty-guided Traversal

DGX agent

arXiv:2606.06721v1 Announce Type: cross Abstract: Robots that operate over extended periods should not merely visit space; they should progressively understand it. Yet most 3D scene graph pipelines tr

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

ShallowBench: Benchmarking Generative Drug Design Models on Shallow-Pocket Targets

DGX agent

arXiv:2606.06717v1 Announce Type: cross Abstract: While generative AI models have demonstrated remarkable success in structure-based drug design, they predominantly rely on deep binding pockets and st

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

Should You Use Your Large Language Model to Explore or Exploit?

DGX agent

arXiv:2502.00225v4 Announce Type: replace-cross Abstract: We evaluate the ability of the current generation of large language models (LLMs) to help a decision-making agent facing an exploration-exploi

agentsarxiv-cs-ai
8 Jun 2026
Research

SleepExplain: Explainable Non-Rapid Eye Movement and Rapid Eye Movement Sleep Stage Classification from EEG Signal

DGX agent

arXiv:2606.07351v1 Announce Type: cross Abstract: Classification of sleep stages is one of the most important diagnostic approaches for a variety of sleep-related disorders. Electroencephalography (EE

researcharxiv-cs-ai
8 Jun 2026
Safety

SlimSearcher: Training Efficiency-Aware Web Agents via Adaptive Reward Gating

DGX agent

arXiv:2606.07074v1 Announce Type: cross Abstract: Deep research agents have demonstrated remarkable capabilities in complex information-seeking tasks, yet this power comes at a steep computational cos

safetyarxiv-cs-ai
8 Jun 2026
Agents

Small Language Model Agents Enable Efficient and High-Quality Knowledge Mining

DGX agent

arXiv:2510.01427v3 Announce Type: replace Abstract: At the core of Deep Research is knowledge mining, the task of extracting structured information from massive unstructured text in response to user i

agentsarxiv-cs-ai
8 Jun 2026
Safety

Socratic-SWE: Self-Evolving Coding Agents via Trace-Derived Agent Skills

DGX agent

arXiv:2606.07412v1 Announce Type: cross Abstract: LLM-driven software engineering agents have become a central testbed for real-world language-model capability, yet their training remains limited by t

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Sparse Subspace-to-Expert Sharing for Task-Agnostic Continual Learning

DGX agent

arXiv:2606.07500v1 Announce Type: cross Abstract: Continual learning in Large Language Models (LLMs) is hindered by the plasticity-stability dilemma, where acquiring new capabilities often leads to ca

model-releasesarxiv-cs-ai
8 Jun 2026
Applications

SpectCount: Spectrotemporal Counting via Synthetic Signals Improves Large Audio Language Models

DGX agent

arXiv:2606.06907v1 Announce Type: cross Abstract: Large audio language models (LALMs) extend large language models with an audio encoder and large-scale audio data. However, the scarcity of high-quali

applicationsarxiv-cs-ai
8 Jun 2026
Tutorials

SS-TPT: Stability and Suitability-Guided Test-Time Prompt Tuning for Adversarially Robust Vision-Language Models

DGX agent

arXiv:2606.06943v1 Announce Type: cross Abstract: Vision-language models (VLMs) such as CLIP achieve strong zero-shot recognition but remain highly fragile under adversarial perturbations. Recent test

tutorialsarxiv-cs-ai
8 Jun 2026
Safety

Stable Reasoning, Unstable Responses: Mitigating LLM Deception via Stability Asymmetry

DGX agent

arXiv:2603.26846v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) expand in capability and application scope, their trustworthiness becomes critical. A vital risk is intrinsic

safetyarxiv-cs-ai
8 Jun 2026
Local Ai

StainFlow: Entity-Stain Tracking and Evidence Linking for Process Rewards in GUI Agents

DGX agent

arXiv:2606.07027v1 Announce Type: new Abstract: Reinforcement Learning (RL) has become a promising approach for improving GUI Agents in long-horizon, stochastic digital environments, but trajectory-le

local-aiarxiv-cs-ai
8 Jun 2026
Applications

Standard vs. Modular Sampling: Best Practices for Reliable LLM Unlearning

DGX agent

arXiv:2509.05316v2 Announce Type: replace-cross Abstract: A conventional LLM Unlearning setting consists of two subsets -'forget' and 'retain', with the objectives of removing the undesired knowledge

applicationsarxiv-cs-ai
8 Jun 2026
Safety

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models

DGX agent

arXiv:2602.02600v3 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have recently emerged as a competitive alternative to autoregressive (AR) models, offering parallel decoding,

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

STREAM: Stochastic Riemannian Flow Matching with Anisotropic Decoder for Digital Histopathology Image Generation

DGX agent

arXiv:2606.07036v1 Announce Type: cross Abstract: Synthetic histopathology image generation addresses critical challenges in computational pathology, including patient privacy and the growing need for

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Superintelligent Retrieval Agent: The Next Frontier of Agentic Retrieval

DGX agent

arXiv:2605.06647v2 Announce Type: replace-cross Abstract: Retrieval-augmented agents are increasingly the interface to large knowledge bases, yet most treat retrieval as a black box: they issue explor

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Supervision versus Demonstration-Based In-Context Learning for Multiword Expression Classification

DGX agent

arXiv:2606.07479v1 Announce Type: cross Abstract: Turkish idiomatic light verb constructions (LVCs) are challenging for multiword expression processing because they often share the same surface form a

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

SV-Detect: AI-generated Text Detection with Steering Vectors

DGX agent

arXiv:2606.07313v1 Announce Type: cross Abstract: Detecting machine-generated text is especially difficult under distribution shift, such as transfer across domains, source models, and editing attacks

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

SW-A^2-Bench: Benchmarking Autonomous Software Agent Generation for Agentic Web

DGX agent

arXiv:2604.04226v2 Announce Type: replace-cross Abstract: The Agentic Web is emerging as a paradigm in which autonomous software agents interact with online resources and with each other to accomplish

model-releasesarxiv-cs-ai
8 Jun 2026
Research

SWE-IF: Aligning Code Evaluation with Human Preference

DGX agent

arXiv:2510.07315v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have catalyzed vibe coding, where users leverage LLMs to generate and iteratively refine code through natural lan

researcharxiv-cs-ai
8 Jun 2026
Research

Synthetic Benchmarks Overstate Forward-Forward Scaling: Real-Data Limits of Layer-Local Training

DGX agent

arXiv:2606.06539v1 Announce Type: cross Abstract: Forward-Forward (FF) learning [Hinton, 2022] replaces backpropagation with strictly layer-local goodness updates. Recent FF-CNN work has narrowed the

researcharxiv-cs-ai
8 Jun 2026
Safety

Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization

DGX agent

arXiv:2606.07000v1 Announce Type: new Abstract: Recent post-training methods, particularly Reinforcement Learning with Verifiable Rewards (RLVR), have significantly enhanced the reasoning ability of L

safetyarxiv-cs-ai
8 Jun 2026
Research

Telling stories, making Hanzi: AI-assisted co-creation with elderly migrants in urban China

DGX agent

arXiv:2507.01548v3 Announce Type: replace-cross Abstract: This paper explores how older migrants in urban China can record stories that everyday language and design often miss. We ran two co-creation

researcharxiv-cs-ai
8 Jun 2026
Model Releases

TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Improved Vision-Language Alignment

DGX agent

arXiv:2606.07451v1 Announce Type: cross Abstract: Vision-language models such as CLIP are highly useful for diverse tasks due to their shared image-text embedding space. Despite this, the image and te

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Textual Supervision Enhances Geospatial Representations in Vision-Language Models

DGX agent

arXiv:2606.07172v1 Announce Type: cross Abstract: Geospatial understanding is a critical yet underexplored dimension in the development of machine learning systems for tasks such as image geolocation

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

The Fine-Tuning Trap: Evaluating Negative Transfer and the Role of PEFT in Sub-1B Mathematical Reasoning

DGX agent

arXiv:2606.06920v1 Announce Type: cross Abstract: Deploying Small Language Models (SLMs) on edge devices requires efficient fine-tuning strategies that adapt models to new tasks without degrading thei

model-releasesarxiv-cs-ai
8 Jun 2026
Local Ai

The Geography of Algorithmic Judgment: LLM Intermediaries, Place Identity, and Racial Steering in Housing Search

DGX agent

arXiv:2606.06694v1 Announce Type: cross Abstract: Large language models (LLMs) are rapidly assuming an intermediary role in housing search through the integration of listing platforms within conversat

local-aiarxiv-cs-ai
8 Jun 2026
Model Releases

The Geometry of Representational Failures in Vision Language Models

DGX agent

arXiv:2602.07025v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) exhibit puzzling failures in multi-object visual tasks, such as hallucinating non-existent elements or failing t

model-releasesarxiv-cs-ai
8 Jun 2026
Research

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook

DGX agent

arXiv:2604.02029v2 Announce Type: replace Abstract: Latent space is rapidly emerging as a native substrate for language-based models. While modern systems are still commonly understood through explici

researcharxiv-cs-ai
8 Jun 2026
Local Ai

The Masked Advantage: Uncovering Local-Language Access to Cultural Knowledge in LLMs

DGX agent

arXiv:2606.07422v1 Announce Type: cross Abstract: Large language models are increasingly used to answer culturally grounded questions across languages, yet it remains unclear whether local cultural kn

local-aiarxiv-cs-ai
8 Jun 2026
Agents

The Sim-to-Real Gap of Foundation Model Agents: A Unified MDP Perspective

DGX agent

arXiv:2606.07017v1 Announce Type: new Abstract: Foundation model agents are increasingly deployed for real-world decision-making, but suffer from the sim-to-real gap. While robotics and classical cont

agentsarxiv-cs-ai
8 Jun 2026
Agents

The Three-Ring Architecture: Governing Agents in the Era of On-Platform Organisations

DGX agent

arXiv:2606.07119v1 Announce Type: cross Abstract: The current phase of enterprise AI deployment faces a structural failure: organisations are acquiring agentic capability without the infrastructure to

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models

DGX agent

arXiv:2606.07157v1 Announce Type: new Abstract: Many efforts to ensure frontier AI models are safe rely on monitoring their chain-of-thought (CoT) reasoning. If models become able to perform sufficien

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Think Like a Pilot: Fine-Grained Long-Horizon UAV Navigation

DGX agent

arXiv:2606.06836v1 Announce Type: cross Abstract: Language-guided UAV agents must execute long-horizon semantic instructions while producing smooth, physically feasible continuous flight commands, yet

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

ThinkBooster: A Unified Framework for Seamless Test-Time Scaling of LLM Reasoning

DGX agent

arXiv:2606.06915v1 Announce Type: cross Abstract: Test-time compute (TTC) scaling has emerged as a powerful paradigm for improving large language model (LLM) reasoning by allocating additional compute

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

TokaMind: A Multi-Modal Transformer Foundation Model for Tokamak Plasma Dynamics

DGX agent

arXiv:2602.15084v2 Announce Type: replace-cross Abstract: We present TokaMind, to our knowledge the first open-source foundation model for tokamak plasma dynamics, based on a Multi-Modal Transformer (

model-releasesarxiv-cs-ai
8 Jun 2026
Research

TOPSIS-RAD: Ranking According to Desires

DGX agent

arXiv:2606.07253v1 Announce Type: new Abstract: Traditional TOPSIS derives its reference points -- the Positive Ideal Solution (PIS) and Negative Ideal Solution (NIS) -- from the observed alternative

researcharxiv-cs-ai
8 Jun 2026
← Previous
1…196197198199200…452
Next →