AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
20,981 results
12 Aug 2026

REDAgentBench: Executable Red Teaming and Faithful Measurement of LLM Agent Systems

Model ReleasesDGX agent

arXiv:2608.10669v1 Announce Type: new Abstract: Large language model (LLM) agents combine language-based reasoning with external tools to perform complex tasks. Adversarial inputs can exploit interact

Reference-Free Post-Training of Open Large Language Models for Multilingual Machine Translation

Model ReleasesDGX agent

arXiv:2608.10812v1 Announce Type: cross Abstract: We study reference-free post-training for multilingual machine translation with open large language models. Starting from the supervised-finetuned MiL

Regression and Classification with Single-Qubit Quantum Neural Networks

ResearchDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2412.09486v2 Announce Type: replace-cross Abstract: The literature reflects a mutually beneficial relationship between machine learning and quantum computing, where progress in one field frequen

Reinforcement Learning-Based Laser Cutting Machine Parameter Optimization

Model ReleasesDGX agent

arXiv:2608.10549v1 Announce Type: new Abstract: Achieving high accuracy in laser-based cutting of optical films requires careful tuning of parameters such as focal length and laser power beam, adjuste

ReLTEx: Reliable LLM-based Taxonomy Expansion

Model ReleasesDGX agent

arXiv:2608.10970v1 Announce Type: cross Abstract: Recent advances in Large Language Models (LLMs) have demonstrated strong capabilities in generating semantically relevant concepts and relations, maki

Rescene: band-limited stochastic forcing turns a frozen neural weather operator into a climate emulator

Model ReleasesDGX agent

arXiv:2608.09971v1 Announce Type: cross Abstract: Over the past few years, the rapid development of machine learning (ML) models for weather forecasting has produced deterministic models whose medium-

Rethinking Text-Based Image Retrieval in Specific Domain

Model ReleasesDGX agent

arXiv:2608.10524v1 Announce Type: cross Abstract: Driven by the rapid advancement of vision-language representation learning, Text-based Image Retrieval (TBIR) has made notable progress. However, exis

Retrieval-Corrected Conformal Prediction for Time Series

Local AiDGX agent

arXiv:2608.10553v1 Announce Type: cross Abstract: Conformal prediction (CP) provides distribution-free prediction intervals for fixed forecasters, but its standard calibration procedure is often ineff

Riemann GeoResolver: A Non-Euclidean Attention Framework from Euclidean Resolver to Hyperbolic-Spherical Geometry

ResearchDGX agent

arXiv:2608.10416v1 Announce Type: cross Abstract: We present a theoretical foundation for inverse-distance attention, from its Euclidean prototype (Resolver) to its non-Euclidean realization (Riemann

RLMOpt: Adaptive Prompt Optimization via Recursive Language Models

Model ReleasesDGX agent

arXiv:2608.10471v1 Announce Type: new Abstract: Prompt optimizers automate the search for prompts that improve language-model performance, but existing methods rely on a predefined optimization proced

Robust Multi-Agent Bandits with Heavy-Tailed Rewards and Information Asymmetry

AgentsDGX agent

arXiv:2608.10529v1 Announce Type: cross Abstract: The multi-armed bandit problem is a central framework in sequential decision-making, extensively studied under sub-Gaussian reward assumptions. Howeve

RTSKG: Building a Rail Transit Station Knowledge Graph Dataset

ResearchDGX agent

arXiv:2608.11080v1 Announce Type: new Abstract: Rail transit systems play a vital role in urban mobility and economic development. As key components of such systems, rail transit stations function as

Rule of Thumb: Explaining Artificial Intelligence Systems using Partial Information

ResearchDGX agent

arXiv:2608.10766v1 Announce Type: new Abstract: Explainable Artificial Intelligence (XAI) seeks to explain how an Artificial Intelligence (AI) system arrived at a particular decision. We propose ''Rul

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning

SafetyDGX agent

arXiv:2608.10513v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) remain vulnerable to jailbreak attacks that exploit visual inputs to bypass safety alignment inherited from their

SBCO: Self-Supervised, Verifier-Grounded Harness Optimization For Planning Agents

SafetyDGX agent

arXiv:2608.10157v1 Announce Type: new Abstract: Self-improving agents seek to reduce the human engineering effort behind AI systems by enabling them to evolve and self-improve their performance over t

Selective Prediction Reduces the Negative Effects of Automation Bias Overall but Increases False Negatives

SafetyDGX agent

arXiv:2508.07617v2 Announce Type: replace-cross Abstract: AI has the potential to augment human decision making. However, even high-performing models can produce inaccurate predictions when deployed.

Self-Correcting Long-Horizon Search Agents via Tree-Structured Memory

ResearchDGX agent

arXiv:2608.10676v1 Announce Type: new Abstract: Large language model (LLM)-based search agents answer questions through multi-step interactions with external environments. However, providing complete

Self-evolving Agentic Customer Support System at LinkedIn

AgentsDGX agent

arXiv:2608.10224v1 Announce Type: new Abstract: Enterprise support agents operate in rapidly changing environments where policies, product capabilities, and knowledge bases evolve continuously, making

Sheaf-Based Federated Representation Learning

Local AiDGX agent

arXiv:2608.10016v1 Announce Type: cross Abstract: Heterogeneous federated systems require agents to learn and exchange informative representations despite differences in data distributions, sensing mo

Similarity Gates Approve Reversals: A Validity Audit of Embedding-Cosine Thresholds in Agent Systems

SafetyDGX agent

arXiv:2608.10216v1 Announce Type: cross Abstract: Agent frameworks ship quality gates that compare text blocks by embedding-cosine similarity and decide at a fixed cutoff. Deduplication filters, seman

Situation Graph Prediction for User Perspective Modeling

Model ReleasesDGX agent

arXiv:2602.13319v2 Announce Type: replace Abstract: Perspective-aware AI requires modeling evolving internal states---goals, emotions, contexts---not merely preferences. Progress is limited by a data

SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models

AgentsDGX agent

arXiv:2608.10538v1 Announce Type: new Abstract: Agent skills represent a standardized format for packaging procedural knowledge and domain expertise, serving within agent harness systems as an essenti

SkillLens: Visual Skill Cards for Retrieval-Augmented GUI Action Prediction and On-Policy Distillation

Model ReleasesDGX agent

arXiv:2608.10775v1 Announce Type: new Abstract: Computer-using agents can perceive rich software interfaces, yet their decisions often lack visual procedural memory: they may recognize individual cont

SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure

ResearchDGX agent

arXiv:2608.11079v1 Announce Type: new Abstract: Self-evolving agents accumulate reusable skills by appending successful procedures and failure fixes. Over time, the same requirement is often restated

sLTN: Structural Logic Tensor Networks

ResearchDGX agent

arXiv:2608.11136v1 Announce Type: new Abstract: Logic Tensor Networks (LTN) provide a neurosymbolic framework in which first-order logic is interpreted through tensor operations, enabling logical cons

Smart Enough to Go Extinct? An Evolutionary Challenge to the Value of General Intelligence and Its Ethical Implications for AGI

SafetyDGX agent

arXiv:2608.10730v1 Announce Type: cross Abstract: The pursuit of artificial general intelligence (AGI) rests on a seemingly self-evident premise: that general intelligence, the kind of flexible, domai

SPIEval: Evaluating Large Language Models as Mobile Assistants over Scattered Personal Information

Model ReleasesDGX agent

arXiv:2608.10692v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as mobile assistants, where a key challenge is leveraging personal information scattered across

SPOTting the Future: Lookahead Explanations for Deep Reinforcement Learning

SafetyDGX agent

arXiv:2608.09967v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) agents achieve strong performance in complex environments, yet their decision-making processes remain difficult to int

Status Association Does Not Reliably Predict Decision Leakage

SafetyDGX agent

arXiv:2608.10089v1 Announce Type: cross Abstract: Bias evaluations often move too quickly from evidence that a model encodes a social association to claims that the same association will alter consequ

Surfacing the Unsaid: CUE-Bench for Affective Stance in Chinese Discourse

Model ReleasesDGX agent

arXiv:2608.10810v1 Announce Type: cross Abstract: Emotion understanding in discourse requires reasoning beyond surface sentiment because speakers often convey affect through indirect, implicit, polite

Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning

SafetyDGX agent

arXiv:2608.11204v1 Announce Type: cross Abstract: Learning reliable surgical manipulation policies is bottlenecked by the scarcity of action-labeled demonstrations: teleoperated surgical robot (e.g.,

SynBoost: A Synergistic Framework for Fast Sampling of Diffusion Models

ResearchDGX agent

arXiv:2506.13058v2 Announce Type: replace-cross Abstract: Diffusion probabilistic models (DPMs) have demonstrated remarkable success in visual generation. However, their iterative sampling mechanism r

TACTICL: Task-Aware Compression of Tabular ICL Models

Model ReleasesDGX agent

arXiv:2608.10837v1 Announce Type: cross Abstract: The strong performance of foundation models for tabular tasks comes at substantial inference costs. Distilling models into task-specific architectures

TAF-MED: Multi-Turn Safety Refusal Collapse in LLMs Under Declared Self-Treatment Intent

Model ReleasesDGX agent

arXiv:2608.10258v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly provide conversational health information that may influence treatment decisions, yet existing benchmarks do

Temporally Grounded Compositional Camera Motion Understanding via Geometric Knowledge Distillation

Model ReleasesDGX agent

arXiv:2608.10932v1 Announce Type: cross Abstract: Understanding camera motion is fundamental to video perception, with applications in spatial intelligence and controllable video generation. Multimoda

Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation

Model ReleasesDGX agent

arXiv:2608.11191v1 Announce Type: cross Abstract: GUI Visual Grounding is a fundamental capability for GUI agents. Existing models typically freeze their parameters after deployment, limiting their ab

The CASE Framework: A Multi-Disciplinary Control Architecture for Governing Enterprise Agentic AI

SafetyDGX agent

arXiv:2608.10153v1 Announce Type: new Abstract: Enterprises are deploying autonomous AI agents faster than they can govern them, and prevailing approaches stretch a single discipline, typically DevSec

The Deliberative Deficit: An Empirical Critique of LLMs in Democratic Discourse

AgentsDGX agent

arXiv:2608.10186v1 Announce Type: cross Abstract: LLMs are increasingly deployed in settings that require collective reasoning on complex, value-laden problems. Confidence in these deployments rests l

The Epistemic Politics of AI Anthropomorphism

ResearchDGX agent

arXiv:2608.00961v2 Announce Type: replace-cross Abstract: AI anthropomorphism is typically treated as a problem of user misperception requiring institutional correction. Users who engage in sustained

The Gaussian-Multinoulli Restricted Boltzmann Machine: A Potts Model Extension of the GRBM

Model ReleasesDGX agent

arXiv:2505.11635v2 Announce Type: cross Abstract: Many real-world tasks, from associative memory to symbolic reasoning, benefit from discrete, structured representations that standard continuous laten

The GenAI Catch-22: Use of Generative Artificial Intelligence in Norwegian Newsrooms During the 2025 Parliamentary Election

ApplicationsDGX agent

arXiv:2608.10773v1 Announce Type: cross Abstract: The increasing use of Generative Artificial Intelligence (GenAI) in journalism raises concerns about possible detrimental effects both on journalism a

The Truth Stays in the Family: Enhancing Contextual Grounding via Inherited Truthful Heads in Model Lineages

Model ReleasesDGX agent

arXiv:2606.15821v2 Announce Type: replace-cross Abstract: Recent advances in large language models (LLMs) have produced many specialized multimodal LLMs (MLLMs) that share common foundational LLMs, fo

ThinkRetrieve: Retrieval-Augmented Reasoning Traces for Test-Time Scaling

TutorialsDGX agent

arXiv:2608.10928v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) improve performance by allocating additional inference-time compute to generate extended chain-of-thought reasoning. Howev

Threat-guided Policy-aware Scene Perturbation for Safe Autonomous Driving with Online Reinforcement Learning

SafetyDGX agent

arXiv:2608.10403v1 Announce Type: new Abstract: Reinforcement learning (RL) has shown promising performance in autonomous driving, yet ensuring the safety of online RL policies remains challenging due

TimeRoute: Time-Aware Modality Routing and Diffusion for Multi-Modal Recommendation

ResearchDGX agent

arXiv:2608.10983v1 Announce Type: cross Abstract: Multi-modal recommenders fuse collaborative signals with item modalities such as text, images, and audio, but the usefulness of each drifts over time

Token-Based Detection of Spurious Correlations in Vision Transformers

Local AiDGX agent

arXiv:2509.04009v2 Announce Type: replace-cross Abstract: Due to their powerful feature association capabilities, neural network-based computer vision models have the ability to detect and exploit uni

Toward a Theory of Value in AI Alignment

SafetyDGX agent

arXiv:2608.10327v1 Announce Type: new Abstract: Can AI systems be aligned to human values? The popularization of large language models (LLMs) and multi-modal foundation models has seen a rise in harms

Toward Human Rights Benchmarking for LLMs: A Pilot Methodology

Model ReleasesDGX agent

arXiv:2608.10268v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate legal determinations over what human rights are realized, and how. Yet, no evaluation benchmark exis

Towards Efficient Reasoning in LLM-Based Recommender Systems via Model Merging

Model ReleasesDGX agent

arXiv:2608.10447v1 Announce Type: cross Abstract: Large language model-based recommender systems are increasingly adopting slow-thinking models that generate step-by-step reasoning before making predi

Towards Sustainable Artificial Intelligence: A Comprehensive Review and Comparative Analysis of Deep Learning Models' Carbon Footprint

ResearchDGX agent

arXiv:2608.09998v1 Announce Type: new Abstract: Artificial Intelligence (AI) and Machine Learning (ML) have become powerful tools for supporting and automating complex human tasks. Despite their benef

Towards Unified Dynamic Face Landmark Detection

Model ReleasesDGX agent

arXiv:2608.10346v1 Announce Type: cross Abstract: Although advancements in face landmark detection (FLD) methods continue to push performance boundaries, they overlook two major functional limitations

TrAC: Trace-Conditioned Answer Consistency for Efficient Uncertainty Quantification in LLMs

ResearchDGX agent

arXiv:2608.00422v2 Announce Type: replace Abstract: Large language models (LLMs) can generate fluent reasoning traces that nevertheless lead to incorrect answers, making response-level uncertainty est

TRACE: Trustworthy Retrieval-Augmented Conversational Engine

Model ReleasesDGX agent

arXiv:2608.10176v1 Announce Type: new Abstract: Public service chatbots are expected to deliver recommendations from an underlying public service directory, while also making sure that the recommendat

TransitReID: Transit OD Data Collection with Occlusion-Resistant Dynamic Passenger Re-Identification

HardwareDGX agent

arXiv:2504.11500v3 Announce Type: replace-cross Abstract: Transit Origin-Destination (OD) data are fundamental for optimizing public transit services, yet current collection methods, such as manual su

Tree-of-Ideas: Automated Research Ideation via Cross-Trajectory Reasoning over Scholarly Evolution

ResearchDGX agent

arXiv:2608.10740v1 Announce Type: new Abstract: Effective research ideation requires moving beyond a static understanding of prior work to trace how research problems and solutions evolve across the l

Two-stage Odd Residual Flows for Mean-Preserving Probabilistic Time Series Forecasting

ResearchDGX agent

arXiv:2608.11114v1 Announce Type: cross Abstract: Probabilistic forecasting plays an essential role in risk-sensitive decision-making, particularly in long-horizon settings. However, existing approach

Uncertainty-Aware Ensemble Deep Randomized Neural Networks for Classification

Model ReleasesDGX agent

arXiv:2608.10007v1 Announce Type: cross Abstract: The current state-of-the-art (SOTA) deep randomized neural networks, such as deep Random Vector Functional Link (dRVFL) and ensemble deep RVFL (edRVFL

Unlocking the Power of Medical Tabular Data via Semantic-Aware Multimodal Pre-training

ResearchDGX agent

arXiv:2608.10522v1 Announce Type: cross Abstract: While vision-language models dominate medical representation learning, unstructured text lacks the dense, quantitative diagnostic phenotypes inherent

Unsupervised Detection of Groundwater Storage Anomalies in Ghana Using GRACE Satellite Data

ResearchDGX agent

arXiv:2608.10233v1 Announce Type: cross Abstract: Groundwater variability in Ghana remains poorly characterized due to limited long-term in-situ observations. This study investigates groundwater stora

UPAIR: Diagnosing Reasoning States via Uncertainty-Progress Alignment for Selective Intervention

SafetyDGX agent

arXiv:2607.17188v2 Announce Type: replace Abstract: While test-time scaling improves the problem-solving ability of large reasoning models (LRMs) through additional inference-time computation, it can

← Previous
123456…350
Next →