AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
26 May 2026

The Impact of Large Language Models on Open-source Innovation: Evidence from GitHub Copilot

ResearchDGX agent

arXiv:2409.08379v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are reshaping knowledge work, yet their impact on voluntary, self-guided open innovation forums (contributors cho

The Many Faces of On-Policy Distillation: Pitfalls, Mechanisms, and Fixes

SafetyDGX agent

arXiv:2605.11182v2 Announce Type: replace Abstract: On-policy distillation (OPD) and on-policy self-distillation (OPSD) have emerged as promising post-training methods for large language models, offer

The Meme Is the Message: Generative Memesis and AI Visuals in the 2024 USA Presidential Elections

ApplicationsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2411.00934v2 Announce Type: replace-cross Abstract: Visual content on social media has become increasingly influential in shaping political discourse and civic engagement, but it also limits par

The Model Is Not the Product: A Dual-Pillar Architecture for Local-First Psychological Coaching

Model ReleasesDGX agent

arXiv:2605.24411v1 Announce Type: new Abstract: Existing language model applications struggle to meet the demand for emotionally oriented support, primarily due to their inability to maintain deep, pe

The Path Matters: Learning a Token-Commitment Policy for Diffusion Language Models

SafetyDGX agent

arXiv:2605.24697v1 Announce Type: cross Abstract: Diffusion large language models promise faster generation by refining many token positions in parallel, but this parallelism introduces a hidden contr

The Time is Here for Just-in-Time Systems: Challenges and Opportunities

Model ReleasesDGX agent

arXiv:2605.24096v1 Announce Type: cross Abstract: Core systems like key-value stores have historically taken years to build, and are designed to be general so as to amortize cost across deployments, p

Theoretical Analysis of Sparse Optimization with Reparameterization, Weight Decay, and Adaptive Learning Rate

ResearchDGX agent

arXiv:2605.25134v1 Announce Type: cross Abstract: Sparse optimization is a fundamental challenge in various practical applications. A popular approach to sparse optimization is ell_p regularization. H

TIAR: Trajectory-Informed Advantage Reweighting for LLM Abstention Learning

Model ReleasesDGX agent

arXiv:2605.25850v1 Announce Type: cross Abstract: This paper investigates large language model (LLM) abstention learning, specifically using ternary reward, which incentivize truthfulness in large lan

TIGER: Text-Informed Generalized Enzyme-Reaction Retrieval

ResearchDGX agent

arXiv:2605.24489v1 Announce Type: new Abstract: Enzyme-reaction retrieval is a fundamental problem in computational biology, underpinning enzyme characterization, reaction mechanism elucidation, and t

Tiny Brains, Giant Impact: Uncovering the Keystone Neurons of LLM with Just a Few Prompts

Model ReleasesDGX agent

arXiv:2605.24846v1 Announce Type: cross Abstract: Large language models (LLMs) display strong comprehensive abilities, yet the internal mechanisms that support these behaviors remain insufficiently un

TinyFormer: Preserving Tiny Objects in YOLO-DETRHybridReal-time Detectors

ResearchDGX agent

arXiv:2605.25046v1 Announce Type: cross Abstract: YOLO-series and DETR-based detectors struggle with tiny-object detection. YOLO-style models benefit from efficient dense prediction, but their large-s

ToolRegistry: A Protocol-Agnostic Tool Management Library for Function-Calling LLMs

Model ReleasesDGX agent

arXiv:2507.10593v3 Announce Type: replace-cross Abstract: Every LLM tool call is structurally an RPC -- a function name, JSON arguments, and a serialized result -- yet each protocol (native Python, MC

TopoAlign: Topology-Aware Visual Representation Alignment

SafetyDGX agent

arXiv:2605.25541v1 Announce Type: cross Abstract: Neural networks encode inputs as high-dimensional vectors, known as representations, that capture how models process data by encoding task-relevant st

Topology-Driven Transferability Estimation of Medical Foundation Models for Segmentation

Model ReleasesDGX agent

arXiv:2602.23916v2 Announce Type: replace-cross Abstract: The advent of large-scale self-supervised learning (SSL) has produced a vast zoo of medical foundation models. However, selecting optimal medi

Toward a Benchmark for Controllable Simulation of Imperfect Students with Large Language Models

Model ReleasesDGX agent

arXiv:2605.25601v1 Announce Type: cross Abstract: Teacher education requires deliberate practice with learners who exhibit identifiable strengths, weaknesses, and partial mastery. Large language model

Toward Enactive Artificial Intelligence

AgentsDGX agent

arXiv:2605.24238v1 Announce Type: new Abstract: In this paper, we advocate for incorporating enactive approaches to perception and cognition into artificial intelligence (AI). Enactive approaches view

Toward Reliable Design of LLM-Enabled Agentic Workflows: Optimizing Latency-Reliability-Cost Tradeoffs

SafetyDGX agent

arXiv:2605.23929v1 Announce Type: new Abstract: Modern AI systems increasingly rely on workflows composed of multiple interacting agents, some powered by large language models (LLMs) and others by con

Towards a Universal Causal Reasoner

ApplicationsDGX agent

arXiv:2605.24873v1 Announce Type: cross Abstract: Despite the importance of causal reasoning, training LLMs to reason causally remains underexplored. Existing data efforts mostly focus on benchmarking

Towards end-to-end LLM-based censoring-aware survival analysis

ResearchDGX agent

arXiv:2605.25399v1 Announce Type: new Abstract: Objective: Survival analysis is central to medical prediction, yet large language models (LLMs) are rarely used as end-to-end survival models because ce

Towards Evaluation Engineering: An Empirical Study of ML Evaluation Harnesses in the Wild

ResearchDGX agent

arXiv:2605.24213v1 Announce Type: cross Abstract: Evaluation harnesses are software systems that orchestrate model evaluation by managing model invocation, data loading, metric computation, and result

Towards Multi-Turn Dialog Systems for Industrial Asset Operations and Maintenance

AgentsDGX agent

arXiv:2605.24953v1 Announce Type: new Abstract: Industrial asset operations and maintenance question answering is inherently multi-turn, iterative, and highly dependent on external tool invocation. Ho

Towards the Connection between Activation Sparsity and Flat Minima

SafetyDGX agent

arXiv:2605.25612v1 Announce Type: cross Abstract: The observation that activation sparsity emerges in MLP blocks of standardly trained Transformers offers an opportunity to drastically reduce computat

Towards trustworthy agentic AI: a comprehensive survey of safety, robustness, privacy, and system security

SafetyDGX agent

arXiv:2605.23989v1 Announce Type: new Abstract: Agentic AI systems -- Large Language Models (LLMs) augmented with planning, tool use, memory, and long-horizon interactions -- can execute complex tasks

TRACER: A Semantic-Aware Framework for Fine-Grained Contamination Detection in Code LLMs

Model ReleasesDGX agent

arXiv:2605.24079v1 Announce Type: cross Abstract: Data contamination is a known threat to the reliability of model evaluation. However, it remains underexplored in code large language models (LLMs), w

TRAFA: Anticipating User Actions to Reduce Errors in Procedural Tasks with Predictive Feedback

ResearchDGX agent

arXiv:2605.24526v1 Announce Type: cross Abstract: Interactive assistance systems typically provide feedback after an action has been completed, supporting error recovery but not preventing the error i

Treatment Effect Estimation with Differentiated Networked Effect on Graph Data

Local AiDGX agent

arXiv:2605.24358v1 Announce Type: cross Abstract: Estimating individual treatment effect (ITE) from observational graph data is crucial for decision-making in the fields such as commerce and medicine.

TriVAL: A Tri-Validation Framework for Faithful Automatic Optimization Modeling

Model ReleasesDGX agent

arXiv:2605.23966v1 Announce Type: cross Abstract: Optimization modeling serves as the pivotal bridge between natural-language problem descriptions and optimization solvers, and remains a cornerstone f

Trust-Aware Joint Feature-Prediction Discrepancy for Robust Domain Adaptation

SafetyDGX agent

arXiv:2605.25119v1 Announce Type: cross Abstract: Domain adaptation aims to mitigate performance degradation caused by distribution shifts between a labeled source domain and an unlabeled or sparsely

Trust but Verify: Prover-Verifier Deliberation for Selective LLM Prediction

Model ReleasesDGX agent

arXiv:2605.25133v1 Announce Type: new Abstract: Reliably knowing when a language model is correct is almost as important as being correct. We introduce prover-verifier deliberation (PVD), an inference

Truthful Online Preference Aggregation for LLM Fine-Tuning in Mobile Crowdsourcing

Model ReleasesDGX agent

arXiv:2605.24052v1 Announce Type: cross Abstract: To better serve users' demands in mobile applications (e.g., navigation), mobile crowdsourcing platforms can iteratively align large language model (L

TS-Skill: A Benchmark for Evaluating Analytical Skills in Time-Series Question Answering

Model ReleasesDGX agent

arXiv:2605.24703v1 Announce Type: cross Abstract: Large language models (LLMs) and time-series language models (TSLMs) are increasingly applied to time-series question answering (TSQA). Unlike text-on

TTPrint: Evidence-Grounded TTP Extraction via Diverge-then-Converge Verification

Model ReleasesDGX agent

arXiv:2605.25836v1 Announce Type: cross Abstract: Extracting MITRE ATT&CK techniques from cyber threat intelligence (CTI) reports is an open-set, multi-label problem requiring both high recall (not mi

Two-Sided Time-Independent Regret for Matching Markets with Limited Interviews

TutorialsDGX agent

arXiv:2602.12224v2 Announce Type: replace-cross Abstract: Two-sided matching platforms rely on preferences from both sides, yet participants can evaluate only a small fraction of potential partners. I

Unbalanced Incomplete Multi-view Clustering via the Scheme of View Evolution: Weak Views are Meat; Strong Views do Eat

ApplicationsDGX agent

arXiv:2011.10254v3 Announce Type: replace-cross Abstract: Incomplete multi-view clustering is an important technique to deal with real-world incomplete multi-view data. Previous works assume that all

Uncertainty Decomposition via Cyclical SG-MCMC and Soft-label Learning for Subjective NLP

Model ReleasesDGX agent

arXiv:2605.24773v1 Announce Type: new Abstract: Annotator disagreement in emotion classification reflects ambiguity intrinsic to emotion concepts and is essential for predictor-quality assessment in s

Uncertainty-DTW for Sequences and Visual Tokens

SafetyDGX agent

arXiv:2605.25110v1 Announce Type: cross Abstract: Aligning structured data is a fundamental problem in computer vision and machine learning, underlying tasks such as time series analysis, human action

Uncertainty Reasoning with Large Language Models for Explainable Disease Diagnosis

TutorialsDGX agent

arXiv:2605.25566v1 Announce Type: new Abstract: Clinical decision-making requires reasoning over incomplete, imprecise, and linguistically expressed patient narratives. While large language models (LL

Uncovering Autoregressive LLM Knowledge of Thematic Fit in Event Representation

ResearchDGX agent

arXiv:2410.15173v4 Announce Type: replace-cross Abstract: The thematic fit estimation task measures semantic arguments' compatibility with a given semantic role for a given predicate. We investigate i

Uncovering Vulnerabilities of LLM-Assisted Cyber Threat Intelligence

ResearchDGX agent

arXiv:2509.23573v4 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to help security analysts manage the surge of cyber threats, automating tasks from vulnerab

Understanding, Accelerating, and Improving MeanFlow Training

ResearchDGX agent

arXiv:2511.19065v2 Announce Type: replace-cross Abstract: MeanFlow promises high-quality generative modeling in few steps, by jointly learning instantaneous and average velocity fields. Yet, the under

Understanding and Mitigating Premature Confidence for Better LLM Reasoning

Model ReleasesDGX agent

arXiv:2605.24396v1 Announce Type: new Abstract: Long chains of thought (CoT) from current language models frequently contain logical gaps and unjustified leaps, limiting the gains from additional test

Understanding Conversational Patterns in Multi-agent Programming: A Case Study on Fibonacci Game Development

Model ReleasesDGX agent

arXiv:2605.24138v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly applied to software engineering (SE), yet their potential for autonomous, role-oriented collaboration re

Uni-DPO: A Unified Paradigm for Dynamic Preference Optimization of LLMs

Model ReleasesDGX agent

arXiv:2506.10054v4 Announce Type: replace-cross Abstract: Direct Preference Optimization (DPO) has emerged as a cornerstone of reinforcement learning from human feedback (RLHF) due to its simplicity a

UniRank: End-to-End Domain-Specific Reranking of Hybrid Text-Image Candidates

SafetyDGX agent

arXiv:2603.29897v2 Announce Type: replace-cross Abstract: Reranking is a critical component in many information retrieval pipelines. Despite remarkable progress in text-only settings, multimodal reran

Unlocking Apple's Private Cloud Compute: An Analysis of Privacy-Preserving Artificial Intelligence

Local AiDGX agent

arXiv:2605.24239v1 Announce Type: cross Abstract: Many existing Artificial Intelligence (AI) solutions on mobile devices rely on an extensive collection of sensitive data, raising privacy concerns and

UtilityMax Prompting: A Formal Framework for Multi-Objective Large Language Model Tasks

Model ReleasesDGX agent

arXiv:2603.11583v4 Announce Type: replace-cross Abstract: The success of a Large Language Model (LLM) task depends heavily on its prompt. Most use-cases specify prompts using natural language, which i

UWM-JEPA: Predictive World Models That Imagine in Belief Space

Model ReleasesDGX agent

arXiv:2605.25313v1 Announce Type: cross Abstract: World models for partially observed environments must imagine multiple compatible hidden futures and steer between them under counterfactual actions.

VaaWIT: Visual-Aware Adaptation of Large Language Models for Multilingual Web Image Translation

Model ReleasesDGX agent

arXiv:2605.24675v1 Announce Type: cross Abstract: Translating text embedded in Web images is crucial for improving content accessibility and cross-lingual information retrieval, particularly within so

vAttention: Verified Sparse Attention

Model ReleasesDGX agent

arXiv:2510.05688v2 Announce Type: replace-cross Abstract: State-of-the-art sparse attention methods for reducing decoding latency fall into two main categories: approximate top-k (and its extension, t

VectorArk: Learning Practical Image Vectorization with Rounded Polygon Representation

ApplicationsDGX agent

arXiv:2605.24398v1 Announce Type: cross Abstract: Recent vision-language model (VLM)-based approaches have achieved impressive results on image vectorization tasks. However, they are typically evaluat

VEN-VL: A Visual Ensemble MoE Framework for Effective and Efficient Multi-Modal Understanding

SafetyDGX agent

arXiv:2605.25952v1 Announce Type: cross Abstract: Despite the remarkable progress achieved by recent efficient methods in accelerating multimodal understanding, they still suffer from noticeable perfo

Verified SHAP: Provable Bounds for Exact Shapley Values of Neural Networks

ResearchDGX agent

arXiv:2605.24084v1 Announce Type: cross Abstract: Shapley additive explanations (SHAP) are widely recognised as computationally intractable for neural networks, since they induce an exponential search

VeriTrace: Evolving Mental Models for Deep Research Agents

Model ReleasesDGX agent

arXiv:2605.26081v1 Announce Type: new Abstract: Deep research agents face vast, interdependent, and pervasively uncertain information. Existing systems explore what evolving intermediate representatio

VineLM: Trie-Based Fine-Grained Control for Agentic Workflows

AgentsDGX agent

arXiv:2605.23914v1 Announce Type: cross Abstract: Agentic workflows interleave configurable LLM stages with tool stages and often include retries or refinement loops. Existing workflow managers profil

VisualOverload: Probing Visual Understanding of VLMs in Really Dense Scenes

Model ReleasesDGX agent

arXiv:2509.25339v3 Announce Type: replace-cross Abstract: Is basic visual understanding really solved in state-of-the-art VLMs? We present VisualOverload, a slightly different visual question answerin

Voting with the Graph: Stable RLAIF via Topological Consistency Maximization

ResearchDGX agent

arXiv:2510.15514v3 Announce Type: replace Abstract: Reinforcement Learning from AI Feedback (RLAIF) relies on LLM judges as preference measurement instruments, yet these instruments are fundamentally

Weakly Supervised Camouflaged Object Detection Based on the SAM Model and Mask Guidance

TutorialsDGX agent

arXiv:2605.25385v1 Announce Type: cross Abstract: Camouflaged object detection (COD) from a single image is a challenging task due to the high similarity between objects and their surroundings. Existi

What Gets Cited: Competitive GEO in AI Answer Engines

SafetyDGX agent

arXiv:2605.25517v1 Announce Type: new Abstract: AI answer engines generate answers from retrieved pages but cite only a few sources. This makes visibility depend not just on ranking, but on being cite

What Happens Next? Anticipating Future Motion by Generating Point Trajectories

ApplicationsDGX agent

arXiv:2509.21592v2 Announce Type: replace-cross Abstract: We consider the problem of forecasting motion from a single image, i.e., predicting how objects in the world are likely to move, without the a

When Can We Trust Early Warnings? Leakage-Excluded Early Outcome Prediction from LMS Interaction Logs

Model ReleasesDGX agent

arXiv:2605.25794v1 Announce Type: new Abstract: Early-warning models built from Learning Management System (LMS) logs aim to predict end-of-course outcomes early enough to enable timely learner suppor

← Previous
1…219220221222223…358
Next →