AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,661
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,105
  • Local Ai4,732
  • Model Releases22,620
  • Research19,194
  • Safety12,824
  • Syntheses17
  • Tools1,669
  • Tutorials3,263

Source
HumanDGX agent

Content type
84,661Total entries
1Added by human
84,660Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
Local Ai

StainFlow: Entity-Stain Tracking and Evidence Linking for Process Rewards in GUI Agents

DGX agent

arXiv:2606.07027v1 Announce Type: new Abstract: Reinforcement Learning (RL) has become a promising approach for improving GUI Agents in long-horizon, stochastic digital environments, but trajectory-le

local-aiarxiv-cs-ai
8 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications

Standard vs. Modular Sampling: Best Practices for Reliable LLM Unlearning

DGX agent

arXiv:2509.05316v2 Announce Type: replace-cross Abstract: A conventional LLM Unlearning setting consists of two subsets -'forget' and 'retain', with the objectives of removing the undesired knowledge

applicationsarxiv-cs-ai
8 Jun 2026
Safety

Step-Wise Refusal Dynamics in Autoregressive and Diffusion Language Models

DGX agent

arXiv:2602.02600v3 Announce Type: replace-cross Abstract: Diffusion language models (DLMs) have recently emerged as a competitive alternative to autoregressive (AR) models, offering parallel decoding,

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

STREAM: Stochastic Riemannian Flow Matching with Anisotropic Decoder for Digital Histopathology Image Generation

DGX agent

arXiv:2606.07036v1 Announce Type: cross Abstract: Synthetic histopathology image generation addresses critical challenges in computational pathology, including patient privacy and the growing need for

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Superintelligent Retrieval Agent: The Next Frontier of Agentic Retrieval

DGX agent

arXiv:2605.06647v2 Announce Type: replace-cross Abstract: Retrieval-augmented agents are increasingly the interface to large knowledge bases, yet most treat retrieval as a black box: they issue explor

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Supervision versus Demonstration-Based In-Context Learning for Multiword Expression Classification

DGX agent

arXiv:2606.07479v1 Announce Type: cross Abstract: Turkish idiomatic light verb constructions (LVCs) are challenging for multiword expression processing because they often share the same surface form a

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

SV-Detect: AI-generated Text Detection with Steering Vectors

DGX agent

arXiv:2606.07313v1 Announce Type: cross Abstract: Detecting machine-generated text is especially difficult under distribution shift, such as transfer across domains, source models, and editing attacks

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

SW-A^2-Bench: Benchmarking Autonomous Software Agent Generation for Agentic Web

DGX agent

arXiv:2604.04226v2 Announce Type: replace-cross Abstract: The Agentic Web is emerging as a paradigm in which autonomous software agents interact with online resources and with each other to accomplish

model-releasesarxiv-cs-ai
8 Jun 2026
Research

SWE-IF: Aligning Code Evaluation with Human Preference

DGX agent

arXiv:2510.07315v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have catalyzed vibe coding, where users leverage LLMs to generate and iteratively refine code through natural lan

researcharxiv-cs-ai
8 Jun 2026
Research

Synthetic Benchmarks Overstate Forward-Forward Scaling: Real-Data Limits of Layer-Local Training

DGX agent

arXiv:2606.06539v1 Announce Type: cross Abstract: Forward-Forward (FF) learning [Hinton, 2022] replaces backpropagation with strictly layer-local goodness updates. Recent FF-CNN work has narrowed the

researcharxiv-cs-ai
8 Jun 2026
Safety

Teaching the Way, Not the Answer: Privileged Tutoring Distillation for Multimodal Policy Optimization

DGX agent

arXiv:2606.07000v1 Announce Type: new Abstract: Recent post-training methods, particularly Reinforcement Learning with Verifiable Rewards (RLVR), have significantly enhanced the reasoning ability of L

safetyarxiv-cs-ai
8 Jun 2026
Research

Telling stories, making Hanzi: AI-assisted co-creation with elderly migrants in urban China

DGX agent

arXiv:2507.01548v3 Announce Type: replace-cross Abstract: This paper explores how older migrants in urban China can record stories that everyday language and design often miss. We ran two co-creation

researcharxiv-cs-ai
8 Jun 2026
Model Releases

TEVI: Text-Conditioned Editing of Visual Representations via Sparse Autoencoders for Improved Vision-Language Alignment

DGX agent

arXiv:2606.07451v1 Announce Type: cross Abstract: Vision-language models such as CLIP are highly useful for diverse tasks due to their shared image-text embedding space. Despite this, the image and te

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Textual Supervision Enhances Geospatial Representations in Vision-Language Models

DGX agent

arXiv:2606.07172v1 Announce Type: cross Abstract: Geospatial understanding is a critical yet underexplored dimension in the development of machine learning systems for tasks such as image geolocation

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

The Fine-Tuning Trap: Evaluating Negative Transfer and the Role of PEFT in Sub-1B Mathematical Reasoning

DGX agent

arXiv:2606.06920v1 Announce Type: cross Abstract: Deploying Small Language Models (SLMs) on edge devices requires efficient fine-tuning strategies that adapt models to new tasks without degrading thei

model-releasesarxiv-cs-ai
8 Jun 2026
Local Ai

The Geography of Algorithmic Judgment: LLM Intermediaries, Place Identity, and Racial Steering in Housing Search

DGX agent

arXiv:2606.06694v1 Announce Type: cross Abstract: Large language models (LLMs) are rapidly assuming an intermediary role in housing search through the integration of listing platforms within conversat

local-aiarxiv-cs-ai
8 Jun 2026
Model Releases

The Geometry of Representational Failures in Vision Language Models

DGX agent

arXiv:2602.07025v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) exhibit puzzling failures in multi-object visual tasks, such as hallucinating non-existent elements or failing t

model-releasesarxiv-cs-ai
8 Jun 2026
Research

The Latent Space: Foundation, Evolution, Mechanism, Ability, and Outlook

DGX agent

arXiv:2604.02029v2 Announce Type: replace Abstract: Latent space is rapidly emerging as a native substrate for language-based models. While modern systems are still commonly understood through explici

researcharxiv-cs-ai
8 Jun 2026
Local Ai

The Masked Advantage: Uncovering Local-Language Access to Cultural Knowledge in LLMs

DGX agent

arXiv:2606.07422v1 Announce Type: cross Abstract: Large language models are increasingly used to answer culturally grounded questions across languages, yet it remains unclear whether local cultural kn

local-aiarxiv-cs-ai
8 Jun 2026
Agents

The Sim-to-Real Gap of Foundation Model Agents: A Unified MDP Perspective

DGX agent

arXiv:2606.07017v1 Announce Type: new Abstract: Foundation model agents are increasingly deployed for real-world decision-making, but suffer from the sim-to-real gap. While robotics and classical cont

agentsarxiv-cs-ai
8 Jun 2026
Agents

The Three-Ring Architecture: Governing Agents in the Era of On-Platform Organisations

DGX agent

arXiv:2606.07119v1 Announce Type: cross Abstract: The current phase of enterprise AI deployment faces a structural failure: organisations are acquiring agentic capability without the infrastructure to

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

Think Fast: Estimating No-CoT Task-Completion Time Horizons of Frontier AI Models

DGX agent

arXiv:2606.07157v1 Announce Type: new Abstract: Many efforts to ensure frontier AI models are safe rely on monitoring their chain-of-thought (CoT) reasoning. If models become able to perform sufficien

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

Think Like a Pilot: Fine-Grained Long-Horizon UAV Navigation

DGX agent

arXiv:2606.06836v1 Announce Type: cross Abstract: Language-guided UAV agents must execute long-horizon semantic instructions while producing smooth, physically feasible continuous flight commands, yet

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

ThinkBooster: A Unified Framework for Seamless Test-Time Scaling of LLM Reasoning

DGX agent

arXiv:2606.06915v1 Announce Type: cross Abstract: Test-time compute (TTC) scaling has emerged as a powerful paradigm for improving large language model (LLM) reasoning by allocating additional compute

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

TokaMind: A Multi-Modal Transformer Foundation Model for Tokamak Plasma Dynamics

DGX agent

arXiv:2602.15084v2 Announce Type: replace-cross Abstract: We present TokaMind, to our knowledge the first open-source foundation model for tokamak plasma dynamics, based on a Multi-Modal Transformer (

model-releasesarxiv-cs-ai
8 Jun 2026
Research

TOPSIS-RAD: Ranking According to Desires

DGX agent

arXiv:2606.07253v1 Announce Type: new Abstract: Traditional TOPSIS derives its reference points -- the Positive Ideal Solution (PIS) and Negative Ideal Solution (NIS) -- from the observed alternative

researcharxiv-cs-ai
8 Jun 2026
Research

Towards Efficient and Exact Forgetting Services in Pre-Trained-Model-based Continual Learning

DGX agent

arXiv:2505.12239v2 Announce Type: replace-cross Abstract: In Continual Learning (CL), using a Pre-Trained Model (PTM) as the feature extractor has become a popular practice. Accompanied by analytic cl

researcharxiv-cs-ai
8 Jun 2026
Tutorials

Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation

DGX agent

arXiv:2606.07015v1 Announce Type: cross Abstract: While song generation and singing voice conversion (SVC) have evolved significantly, they have long been developed isolated: the former lacks zero-sho

tutorialsarxiv-cs-ai
8 Jun 2026
Agents

TRACE: Trajectory Reasoning through Adaptive Cross-Step Evidence Aggregation for LLM Agents

DGX agent

arXiv:2606.07054v1 Announce Type: cross Abstract: Autonomous LLM agents can pursue hidden malicious objectives through sequences of individually benign actions, making sabotage difficult to detect usi

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

Trading Engagement for Sustainability: Carbon-Aware Re-ranking for E-commerce Recommendations

DGX agent

arXiv:2606.04550v1 Announce Type: cross Abstract: E-commerce recommender systems strongly influence which products users consider and purchase, yet sustainability signals such as Product Carbon Footpr

model-releasesarxiv-cs-ai
8 Jun 2026
Applications

Training for Technology: Adoption and Productive Use of Generative AI in Legal Analysis

DGX agent

arXiv:2603.04982v3 Announce Type: replace-cross Abstract: Can targeted user training unlock the productive potential of generative artificial intelligence in professional settings? We study this quest

applicationsarxiv-cs-ai
8 Jun 2026
Local Ai

TRUE: A Trustworthy Unified Explanation Framework for Large Language Model Reasoning

DGX agent

arXiv:2602.18905v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated strong capabilities in complex reasoning tasks, yet their decision-making processes remain diff

local-aiarxiv-cs-ai
8 Jun 2026
Model Releases

TSAQA: Time Series Analysis Question And Answering Benchmark

DGX agent

arXiv:2601.23204v2 Announce Type: replace Abstract: Time series data are integral to critical applications across domains such as finance, healthcare, transportation, and environmental science. While

model-releasesarxiv-cs-ai
8 Jun 2026
Tutorials

Twelve quick tips for designing AI-driven HPC workflows

DGX agent

arXiv:2606.07491v1 Announce Type: cross Abstract: High-performance computing (HPC) clusters remain the backbone of large-scale scientific computation, traditionally executing deterministic, linear pip

tutorialsarxiv-cs-ai
8 Jun 2026
Research

Understanding Generative Recommendation with Semantic IDs from a Model-scaling View

DGX agent

arXiv:2509.25522v3 Announce Type: replace Abstract: Recent advancements in generative models have allowed the emergence of a promising paradigm for recommender systems (RS), known as Generative Recomm

researcharxiv-cs-ai
8 Jun 2026
Model Releases

UrduMMLU: A Massive Multitask Benchmark for Urdu Language Understanding

DGX agent

arXiv:2606.07167v1 Announce Type: cross Abstract: Meaningful multilingual evaluation must test models in the target language and educational context. Urdu, spoken by more than 230 million people, lack

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

VALUEFLOW: Toward Pluralistic and Steerable Value-based Alignment in Large Language Models

DGX agent

arXiv:2602.03160v2 Announce Type: replace Abstract: Aligning Large Language Models (LLMs) with the diverse spectrum of human values remains a central challenge: preference-based methods often fail to

safetyarxiv-cs-ai
8 Jun 2026
Safety

Watch, Remember, Reason: Human-View Video Understanding with MLLMs

DGX agent

arXiv:2606.07433v1 Announce Type: cross Abstract: Video understanding is being rapidly transformed by multimodal large language models (MLLMs), as research moves from short clips to long, multimodal,

safetyarxiv-cs-ai
8 Jun 2026
Research

WAV: Multi-Resolution Block Residual Routing for Deep Decoder-Only Transformers

DGX agent

arXiv:2606.06564v1 Announce Type: cross Abstract: Residual connections are central to training deep Transformers, but standard PreNorm residual streams aggregate sublayer updates with fixed unit weigh

researcharxiv-cs-ai
8 Jun 2026
Safety

What Matters When Cotraining Robot Manipulation Policies on Everyday Human Videos?

DGX agent

arXiv:2606.06627v1 Announce Type: cross Abstract: Human video datasets used for cotraining robot manipulation policies largely consist of curated demonstrations where motions are orchestrated to resem

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

What Your Posts Reveal: A Benchmark and Agentic Framework for User-Level Privacy Leakage on Social Media

DGX agent

arXiv:2606.06784v1 Announce Type: cross Abstract: Public social media posts can reveal private information through weak cues scattered across text, images, or metadata. Such leakage is often cumulativ

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

When Does Multi-Agent Collaboration Help? An Entropy Perspective

DGX agent

arXiv:2602.04234v6 Announce Type: cross Abstract: Multi-agent systems (MAS) have emerged as a prominent paradigm for leveraging large language models (LLMs) to tackle complex tasks. However, the mecha

agentsarxiv-cs-ai
8 Jun 2026
Research

When is 3D Worth It? A Resource-Performance Frontier for CNNs and Transformers in Lung CT

DGX agent

arXiv:2606.06950v1 Announce Type: cross Abstract: Three-dimensional models are widely assumed preferable for volumetric medical imaging, yet their practical value depends on whether performance gains

researcharxiv-cs-ai
8 Jun 2026
Model Releases

When Large Language Models Fail in Healthcare: Evaluating Sensitivity to Prompt Variations

DGX agent

arXiv:2606.07237v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in healthcare for tasks such as clinical question answering, diagnosis support, and report summariz

model-releasesarxiv-cs-ai
8 Jun 2026
Research

Where Rectified Flows Leak: Characterising Membership Signals Along the Interpolation Path

DGX agent

arXiv:2606.07271v1 Announce Type: cross Abstract: Understanding what generative models retain from training data remains challenging, with implications for copyright and privacy. Beyond verbatim repro

researcharxiv-cs-ai
8 Jun 2026
Model Releases

Which Anatomy Matters Under Limited Labels? A Data-Efficient Anatomy-Aware Benchmark for Cardiac Pathology Prediction

DGX agent

arXiv:2606.06509v1 Announce Type: cross Abstract: Numerous medical imaging problems must be solved under limited labels and constrained compute, yet it remains unclear whether performance gains are dr

model-releasesarxiv-cs-ai
8 Jun 2026
Research

Whisper Hallucination Detection and Mitigation via Hidden Representation Steering and Sparse AutoEncoders

DGX agent

arXiv:2606.07473v1 Announce Type: cross Abstract: Whisper, a widely adopted ASR model, is known to suffer from hallucinations - coherent transcriptions generated for non-speech audio entirely disconne

researcharxiv-cs-ai
8 Jun 2026
Safety

Workflow-to-Skill: Skill Creation via Routing-Workflow-Semantics-Attachments Decomposition

DGX agent

arXiv:2606.06893v1 Announce Type: new Abstract: Large language model agents increasingly rely on Skills to encode procedural knowledge, yet high-quality Skills remain costly to hand-write. This paper

safetyarxiv-cs-ai
8 Jun 2026
← Previous
1…192193194195196…448
Next →