AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlog
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,236 results
Local Ai

Token-Based Detection of Spurious Correlations in Vision Transformers

DGX agent

arXiv:2509.04009v2 Announce Type: replace-cross Abstract: Due to their powerful feature association capabilities, neural network-based computer vision models have the ability to detect and exploit uni

local-aiarxiv-cs-ai
12 Aug 2026
Safety

Toward a Theory of Value in AI Alignment

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2608.10327v1 Announce Type: new Abstract: Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms

safetyarxiv-cs-ai
12 Aug 2026
Model Releases

Toward Human Rights Benchmarking for LLMs: A Pilot Methodology

DGX agent

arXiv:2608.10268v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate legal determinations over what human rights are realized, and how. Yet, no evaluation benchmark exis

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Towards Efficient Reasoning in LLM-Based Recommender Systems via Model Merging

DGX agent

arXiv:2608.10447v1 Announce Type: cross Abstract: Large language model-based recommender systems are increasingly adopting slow-thinking models that generate step-by-step reasoning before making predi

model-releasesarxiv-cs-ai
12 Aug 2026
Research

Towards Sustainable Artificial Intelligence: A Comprehensive Review and Comparative Analysis of Deep Learning Models' Carbon Footprint

DGX agent

arXiv:2608.09998v1 Announce Type: new Abstract: Artificial Intelligence (AI) and Machine Learning (ML) have become powerful tools for supporting and automating complex human tasks. Despite their benef

researcharxiv-cs-ai
12 Aug 2026
Model Releases

Towards Unified Dynamic Face Landmark Detection

DGX agent

arXiv:2608.10346v1 Announce Type: cross Abstract: Although advancements in face landmark detection (FLD) methods continue to push performance boundaries, they overlook two major functional limitations

model-releasesarxiv-cs-ai
12 Aug 2026
Research

TrAC: Trace-Conditioned Answer Consistency for Efficient Uncertainty Quantification in LLMs

DGX agent

arXiv:2608.00422v2 Announce Type: replace Abstract: Large language models (LLMs) can generate fluent reasoning traces that nevertheless lead to incorrect answers, making response-level uncertainty est

researcharxiv-cs-ai
12 Aug 2026
Model Releases

TRACE: Trustworthy Retrieval-Augmented Conversational Engine

DGX agent

arXiv:2608.10176v1 Announce Type: new Abstract: Public service chatbots are expected to deliver recommendations from an underlying public service directory, while also making sure that the recommendat

model-releasesarxiv-cs-ai
12 Aug 2026
Hardware

TransitReID: Transit OD Data Collection with Occlusion-Resistant Dynamic Passenger Re-Identification

DGX agent

arXiv:2504.11500v3 Announce Type: replace-cross Abstract: Transit Origin-Destination (OD) data are fundamental for optimizing public transit services, yet current collection methods, such as manual su

hardwarearxiv-cs-ai
12 Aug 2026
Research

Tree-of-Ideas: Automated Research Ideation via Cross-Trajectory Reasoning over Scholarly Evolution

DGX agent

arXiv:2608.10740v1 Announce Type: new Abstract: Effective research ideation requires moving beyond a static understanding of prior work to trace how research problems and solutions evolve across the l

researcharxiv-cs-ai
12 Aug 2026
Research

Two-stage Odd Residual Flows for Mean-Preserving Probabilistic Time Series Forecasting

DGX agent

arXiv:2608.11114v1 Announce Type: cross Abstract: Probabilistic forecasting plays an essential role in risk-sensitive decision-making, particularly in long-horizon settings. However, existing approach

researcharxiv-cs-ai
12 Aug 2026
Model Releases

Uncertainty-Aware Ensemble Deep Randomized Neural Networks for Classification

DGX agent

arXiv:2608.10007v1 Announce Type: cross Abstract: The current state-of-the-art (SOTA) deep randomized neural networks, such as deep Random Vector Functional Link (dRVFL) and ensemble deep RVFL (edRVFL

model-releasesarxiv-cs-ai
12 Aug 2026
Research

Unlocking the Power of Medical Tabular Data via Semantic-Aware Multimodal Pre-training

DGX agent

arXiv:2608.10522v1 Announce Type: cross Abstract: While vision-language models dominate medical representation learning, unstructured text lacks the dense, quantitative diagnostic phenotypes inherent

researcharxiv-cs-ai
12 Aug 2026
Research

Unsupervised Detection of Groundwater Storage Anomalies in Ghana Using GRACE Satellite Data

DGX agent

arXiv:2608.10233v1 Announce Type: cross Abstract: Groundwater variability in Ghana remains poorly characterized due to limited long-term in-situ observations. This study investigates groundwater stora

researcharxiv-cs-ai
12 Aug 2026
Safety

UPAIR: Diagnosing Reasoning States via Uncertainty-Progress Alignment for Selective Intervention

DGX agent

arXiv:2607.17188v2 Announce Type: replace Abstract: While test-time scaling improves the problem-solving ability of large reasoning models (LRMs) through additional inference-time computation, it can

safetyarxiv-cs-ai
12 Aug 2026
Model Releases

UserToolBench: A User-Profile-Hidden Benchmark for Personalized Decision Making in Tool-Use LLMs

DGX agent

arXiv:2608.10042v1 Announce Type: cross Abstract: Tool-use LLMs are increasingly asked to act on users' behalf, but existing benchmarks usually focus on profile recall, style imitation, generic tool u

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

V-FiLLM: Verified Financial LLM Reasoning Benchmark

DGX agent

arXiv:2608.11047v1 Announce Type: new Abstract: While existing benchmarks have made substantial progress in evaluating LLMs across STEM domains, financial reasoning over structured data remains compar

model-releasesarxiv-cs-ai
12 Aug 2026
Agents

VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection

DGX agent

arXiv:2511.19436v2 Announce Type: replace-cross Abstract: Existing Video Detailed Captioning (VDC) methods predominantly rely on costly human annotations or distillation from powerful proprietary mode

agentsarxiv-cs-ai
12 Aug 2026
Research

VERDICT: Training-Free Step-Wise Verification of Multimodal Reasoning via Disagreement-Aware Consensus

DGX agent

arXiv:2608.10665v1 Announce Type: new Abstract: Multimodal large language models often generate reasoning chains containing subtle errors that lead to incorrect answers. Current verification approache

researcharxiv-cs-ai
12 Aug 2026
Model Releases

VibeLifeBench: Can Your Life Agent Be Proactive and Persistent in a Living World?

DGX agent

arXiv:2608.10875v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly deployed as personal assistants. Existing evaluations, however, mostly use short, self-contained re

model-releasesarxiv-cs-ai
12 Aug 2026
Safety

What We Know about Responsible AI Practices in Industry: A Half Decade of Empirical Research

DGX agent

arXiv:2608.10431v1 Announce Type: cross Abstract: Responsible AI (RAI) has become a central concern for technology companies, regulators, and the public. How industry practitioners interpret, implemen

safetyarxiv-cs-ai
12 Aug 2026
Agents

When Agent Automation Becomes Profitable: Quantifying and Insuring Autonomous AI Risk through Trace-Economic Underwriting

DGX agent

arXiv:2606.16465v2 Announce Type: replace Abstract: AI agents can now take irreversible actions in operational systems, but agent-caused losses are still not clearly assigned, priced, or transferred.

agentsarxiv-cs-ai
12 Aug 2026
Model Releases

When Chain-of-Thought Helps and When It Hurts: An Empirical Investigation of the Serial-Depth Bottleneck in LLM Reasoning

DGX agent

arXiv:2608.09942v1 Announce Type: cross Abstract: It is widely assumed that chain-of-thought (CoT) prompting universally improves LLM reasoning. We investigate this through the conceptual framework of

model-releasesarxiv-cs-ai
12 Aug 2026
Research

Whisper-Aware LLM: Self-Supervised Uncertainty Learning for Robust Whispered Speech Recognition

DGX agent

arXiv:2608.10836v1 Announce Type: cross Abstract: The signal ambiguity of whispered speech drives ASR systems toward two opposing failure modes: failing to capture whispered speech or hallucinatory tr

researcharxiv-cs-ai
12 Aug 2026
Model Releases

Why Does CLAUDE.md Keep Growing? Catastrophic Remembering in Agentic Coding

DGX agent

arXiv:2608.11095v1 Announce Type: new Abstract: Agentic coding READMEs like CLAUDE.md grow without bound in real repositories, stopping only when the repository retires or someone rewrites the file wh

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Withholding the Completing Chunk: Deterministic Pair-Completion Guardrails for Streaming LLM Output

DGX agent

arXiv:2608.10279v1 Announce Type: cross Abstract: Streaming language-model output creates a release-timing problem: complete-response moderation acts after streamed text has escaped, whereas repeated

model-releasesarxiv-cs-ai
12 Aug 2026
Model Releases

Workflow Cards: Structured Summaries of Workflow Executions Using Provenance Data

DGX agent

arXiv:2608.11022v1 Announce Type: cross Abstract: Model Cards and Data Cards have demonstrated the value of structured, human-readable documentation for machine learning artifacts, capturing their con

model-releasesarxiv-cs-ai
12 Aug 2026
Safety

XCoT-VLA: Executable Chain-of-Thought for Vision-Language-Action Driving

DGX agent

arXiv:2608.10976v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models can connect scene understanding, semantic reasoning, and trajectory generation for autonomous driving. However, verb

safetyarxiv-cs-ai
12 Aug 2026
Research

'YES! YES! I absolutely love this insight!' Affirmative Narration as Interactional Strategy in Dialogues with LLM Chatbots

DGX agent

arXiv:2607.28646v2 Announce Type: cross Abstract: This article analyses narrative mechanisms that are common in dialogues with LLM chatbots. In combination, these mechanisms produce an interactional s

researcharxiv-cs-ai
12 Aug 2026
Safety

Your LLM, Your Style: Behavioral Mode Axes for LLM Behavioral Control

DGX agent

arXiv:2608.10703v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly act in interactive settings where their behavioral styles affect user experience, safety, and downstream dec

safetyarxiv-cs-ai
12 Aug 2026
Model Releases

360CityArena: A Realistic Virtual Urban Navigation Benchmark for Embodied Agents

DGX agent

arXiv:2608.08814v1 Announce Type: cross Abstract: We present 360CityArena, a benchmark for evaluating the urban exploration capabilities of embodied agents within a photorealistic environment construc

model-releasesarxiv-cs-ai
11 Aug 2026
Research

A Combined Feature-Based Framework for Disguise and Spoofing Detection in Face Recognition Systems

DGX agent

arXiv:2608.08521v1 Announce Type: cross Abstract: Face recognition systems face two distinct, commonly-separated failure modes: spoofing, where an impostor presents a photograph or video of an authori

researcharxiv-cs-ai
11 Aug 2026
Model Releases

A Fair Objective for Human-Empowerment-Preserving AI: Desiderata, Design, and Likely Behavioral Consequences

DGX agent

arXiv:2608.08240v1 Announce Type: new Abstract: This paper explores the idea of promoting well-being and safety in human-AI interactions by forcing AI agents explicitly to empower humans and to manage

model-releasesarxiv-cs-ai
11 Aug 2026
Research

A foundation model of numerical intelligence with cross-disciplinary generalization

DGX agent

arXiv:2607.28432v2 Announce Type: replace Abstract: Intelligence is commonly understood as the ability to acquire and apply knowledge, adapt to unfamiliar situations and solve new problems. Large lang

researcharxiv-cs-ai
11 Aug 2026
Agents

A Grounded and Decomposed Framework for Relation-Level Hallucination Evaluation in Abstractive Summarization

DGX agent

arXiv:2608.08180v1 Announce Type: cross Abstract: Abstractive text summarization systems frequently generate fluent yet unfaithful summaries by fabricating or distorting relationships between entities

agentsarxiv-cs-ai
11 Aug 2026
Research

A Minimal kappa--au Logic for Risk-Sensitive Abduction

DGX agent

arXiv:2608.08192v1 Announce Type: new Abstract: Standard approaches to abductive reasoning can retain multiple candidate explanations, but they do not generally combine explicit compositional cross-hy

researcharxiv-cs-ai
11 Aug 2026
Research

A Multi-Scale Temporal Framework with Dynamic Fusion for EEG-Based Emotion Recognition

DGX agent

arXiv:2608.09088v1 Announce Type: new Abstract: Mixed emotions represent a clinically relevant but still underexplored target for automatic emotion recognition. EEG provides millisecond-level access t

researcharxiv-cs-ai
11 Aug 2026
Research

A New Approach to Characterising Optimisation Problems Using Programmatic Representation and Complexity Measures

DGX agent

arXiv:2608.08898v1 Announce Type: cross Abstract: Characterising optimisation problem instances is a fundamental part of understanding the behaviour and performance of different algorithms as well as

researcharxiv-cs-ai
11 Aug 2026
Research

A QUBO-Inspired Computational Framework for Airport Landside Bottleneck Diagnosis and Dynamic Dispatch Optimization

DGX agent

arXiv:2608.08632v1 Announce Type: new Abstract: Airport landside traffic centers connect terminal arrivals with taxis, ride-hailing vehicles, private cars, buses, metro services, parking facilities, a

researcharxiv-cs-ai
11 Aug 2026
Model Releases

A Rigorous Turing Test: a Foundation for Evaluating Artificial General Intelligence

DGX agent

arXiv:2501.17629v2 Announce Type: replace-cross Abstract: Several studies claim that large language models have passed the Turing Test and hence can 'think', yet none follow Turing's original instruct

model-releasesarxiv-cs-ai
11 Aug 2026
Research

A Sobering Look at Tabular Data Generation via Probabilistic Circuits

DGX agent

arXiv:2603.23016v2 Announce Type: replace-cross Abstract: Tabular data is more challenging to generate than text and images, due to its heterogeneous features and much lower sample sizes. On this task

researcharxiv-cs-ai
11 Aug 2026
Safety

A Structural Dynamics Graph World Model: Unified Modeling, Constrained Rollout, and Interpretable Calibration

DGX agent

arXiv:2608.08689v1 Announce Type: new Abstract: The state evolution of a complex system arises jointly from object laws, relational propagation, domain conservation, and unmodeled error. Forcing all s

safetyarxiv-cs-ai
11 Aug 2026
Safety

A Unified Framework for Dynamic Reward Shaping in Reinforcement Learning

DGX agent

arXiv:2608.08158v1 Announce Type: new Abstract: Sparse, delayed, and weakly informative rewards remain central obstacles to efficient reinforcement learning. Reward shaping addresses these limitations

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

A Unified Issue Resolution Benchmark for Requirement Clarification, Planning, and Code Generation for Coding Agents

DGX agent

arXiv:2608.09072v1 Announce Type: cross Abstract: Large language model-powered coding agents are increasingly used to modify existing code repositories, for example, by adding features or fixing bugs.

model-releasesarxiv-cs-ai
11 Aug 2026
Research

Abstracted Away: Resisting Alienation and Ungrounded Abstraction in AI Research Communities

DGX agent

arXiv:2608.08408v1 Announce Type: cross Abstract: Logics of abstraction in computational AI research often push important forms of knowledge and reflection aside: dominant standards of legitimacy sepa

researcharxiv-cs-ai
11 Aug 2026
Model Releases

ACEvo: Adversarial Co-Evolution of Problem Distributions and Solvers for Combinatorial Optimization

DGX agent

arXiv:2506.02594v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used to synthesize heuristic programs, yet most existing pipelines optimize solvers against fixed benc

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

ActBench: Self-Evolving Benchmark of Behavioral Safety in Cowork Agents

DGX agent

arXiv:2608.09476v1 Announce Type: cross Abstract: Cowork agents may complete benign tasks while disclosing protected data, manipulating unauthorized state, invocate unauthorized API. We define behavio

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

ActiveFly-Bench: Aligning Embodied Question Answering with Vision-Language-Action for Aerial Embodied Perception

DGX agent

arXiv:2607.10180v2 Announce Type: replace-cross Abstract: We introduce ActiveFly-Bench, the first benchmark to bridge cyberspace reasoning and physical-world interaction for UAV embodied perception. T

model-releasesarxiv-cs-ai
11 Aug 2026
← Previous
1…910111213…443
Next →