AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
9 Jun 2026

Language-based Trial and Error Falls Behind in the Era of Experience

Model ReleasesDGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

Larch: Learned Query Optimization for Semantic Predicates

ApplicationsDGX agent

arXiv:2606.07923v1 Announce Type: cross Abstract: With the advent of Large Language Models (LLMs), many database systems introduced semantic operators that enabled analytical queries over unstructured

Large Language Models for Imbalanced Classification: Diversity makes the difference

ResearchDGX agent

arXiv:2510.09783v2 Announce Type: replace-cross Abstract: Oversampling is one of the most widely used approaches for addressing imbalanced classification. The core idea is to generate additional minor


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Large Language Models Should Learn Personalized Rather Than Aggregated Human Preferences

SafetyDGX agent

arXiv:2606.07629v1 Announce Type: cross Abstract: Current approaches to aligning large language models (LLMs) aggregate diverse human preferences into a single reward signal, effectively optimizing fo

Large Models for Time Series and Spatio-Temporal Data: A Survey and Outlook

ApplicationsDGX agent

arXiv:2310.10196v3 Announce Type: replace-cross Abstract: Temporal data, including time series and spatio-temporal data, are pervasive in real-world applications. Generated in massive volumes by physi

LargeMonitor: Monitoring Online Task-Free Continual Learning via Large Pretrained Models

Model ReleasesDGX agent

arXiv:2606.09430v1 Announce Type: cross Abstract: Online task-free continual learning (TFCL) requires intelligent agents to sequentially accumulate knowledge from an unbounded, non-stationary data str

Late-Layer Fusion is Enough: Dual-Path Vision Token Routing for Multimodal Large Language Models under Visual Saturation

ResearchDGX agent

arXiv:2606.09131v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) commonly inherit the deep, symmetric Transformer backbone designed for unimodal text modeling, and apply the sa

Latent Diffusion Policy: Shaping Latent Spaces for Diffusion-Based Robotic Manipulation

SafetyDGX agent

arXiv:2606.08657v1 Announce Type: cross Abstract: Diffusion-based visuomotor policies operating directly in raw action spaces conflate scene comprehension with trajectory generation within a single de

LATTEArena: An Evaluation Framework for LLM-powered Tabular Feature Engineering (Extended Version)

ResearchDGX agent

arXiv:2606.09004v1 Announce Type: new Abstract: Feature engineering remains essential for tabular data analysis, and Large Language Models (LLMs) have emerged as a promising paradigm for automating th

LCAM: A Framework for Diagnosing Interactional Alignment Failures in Con-versational AI

SafetyDGX agent

arXiv:2606.08131v1 Announce Type: cross Abstract: Conversational AI is increasingly used for advice, interpretation, reassurance, and decision support in contexts where users may be vulnerable, uncert

LEAF: Growing Trees Without Branching for Speech-Aware Large Language Model Post-Training

Model ReleasesDGX agent

arXiv:2606.07610v1 Announce Type: cross Abstract: State-of-the-art GRPO-style methods for speech-aware large language model post-training suffer from coarse credit assignment, broadcasting the same te

Learning Quantized Continuous Controllers for Integer Hardware

SafetyDGX agent

arXiv:2511.07046v4 Announce Type: replace-cross Abstract: Deploying continuous-control reinforcement learning policies on embedded hardware requires meeting tight latency and power budgets. Small FPGA

Learning to Attack and Defend: Adaptive Red Teaming of Language Models via GRPO

SafetyDGX agent

arXiv:2606.09701v1 Announce Type: cross Abstract: AI red teaming must continually adapt to evolving attackers and defenders. Reinforcement learning offers a promising approach to discovering novel att

Leveraging Structural Constraints for Diffusion-based Neural TSP Solvers

Local AiDGX agent

arXiv:2606.09343v1 Announce Type: new Abstract: Neural combinatorial optimization has recently achieved strong results on the Euclidean Traveling Salesman Problem (TSP) using generative models such as

LFNO: Bridging Laplace and Fourier via Transient-Steady Decomposition

ResearchDGX agent

arXiv:2606.07601v1 Announce Type: cross Abstract: We introduce the Laplace-Fourier Neural Operator (LFNO), a unified framework for modeling dynamical systems across transient and steady-state regimes

Liberating LLM Capabilities in Full-Duplex Speech Models

ResearchDGX agent

arXiv:2606.07547v1 Announce Type: cross Abstract: Speech-based large language models are typically constrained to spoken replies, which limits their user-facing outputs to what can be verbalized and s

Liquid Neural Networks as a Drop-in Continuous-Time Deformation Field for Dynamic 3D Gaussian Splatting

ResearchDGX agent

arXiv:2606.07670v1 Announce Type: cross Abstract: Deformable 3D Gaussian Splatting (D-3DGS) re-constructs dynamic scenes from monocular video by deforming a canonical set of 3D Gaussians through a pos

LLM-Orchestrated Conformance Checking in Stroke Care Without Computer-Interpretable Guidelines

ApplicationsDGX agent

arXiv:2606.09489v1 Announce Type: new Abstract: Objective: Conformance checking in healthcare seeks to assess whether patient care pathways adhere to clinical guidelines. However, its practical applic

LogNEO: A GPT-Neo Reinforcement Learning Framework for Accurate Real-Time Log Anomaly Detection

SafetyDGX agent

arXiv:2606.08153v1 Announce Type: cross Abstract: Detecting anomalies in large-scale system logs is critical for the reliability and security of modern computing infrastructure. We present LogNEO, a l

Lost in the Flow with Code Talkers: Unveiling the Instruction-Tuning Tax of Large Language Models in Code Tasks

ResearchDGX agent

arXiv:2606.08676v1 Announce Type: cross Abstract: AI coding assistants have significantly improved developer productivity by automatically suggesting code that aligns with user intent, and many of the

LoTUS: Large-Scale Machine Unlearning with a Taste of Uncertainty

ApplicationsDGX agent

arXiv:2503.18314v5 Announce Type: replace-cross Abstract: We present LoTUS, a novel Machine Unlearning (MU) method that eliminates the influence of training samples from pre-trained models, avoiding r

MAR:Multi-Agent Reflexion Improves Reasoning Abilities in LLMs

AgentsDGX agent

arXiv:2512.20845v2 Announce Type: replace Abstract: LLMs have shown the capacity to improve their performance on reasoning tasks through reflecting on their mistakes, and acting with these reflections

MASS: Deep Research for Social Sciences with Memory-Augmented Social Simulation

AgentsDGX agent

arXiv:2606.09198v1 Announce Type: new Abstract: Deep Research agents powered by Large Language Models (LLMs) have exhibited extraordinary potential in automated paper writing tasks. However, existing

MatMind: A Structure-Activity Knowledge-Driven Generative Foundation Model for Materials Science

ResearchDGX agent

arXiv:2606.07712v1 Announce Type: cross Abstract: Progress in AI-driven crystal materials science has so far been carried by narrow architectures purpose-built for individual tasks -- graph neural net

MatSciBench: Benchmarking the Reasoning Ability of Large Language Models in Materials Science

Model ReleasesDGX agent

arXiv:2510.12171v2 Announce Type: replace Abstract: Large Language Models have shown strong scientific reasoning ability, but their performance on materials science problems remains less studied. To f

MBABench: Evaluating LLM Agents on End-to-End Spreadsheet Tasks in Finance

Model ReleasesDGX agent

arXiv:2605.22664v2 Announce Type: replace Abstract: LLM agents are increasingly expected to carry out end-to-end workflows, producing complete artifacts from high-level user instructions. To meet ente

MC-CPO: Mastery-Conditioned Constrained Policy Optimization for Pedagogically Safe Intelligent Tutoring Systems

SafetyDGX agent

arXiv:2604.04251v2 Announce Type: replace Abstract: Intelligent tutoring systems increasingly rely on reinforcement learning to personalise instruction, yet optimising for observable engagement signal

MC-PDD: Masked Corpus-Level Pretraining Data Detection for Black-Box Large Language Models

SafetyDGX agent

arXiv:2606.07996v1 Announce Type: cross Abstract: Pretraining is fundamental to the development of Large Language Models (LLMs), yet the opacity of pretraining data complicates model analysis and rais

Mechanistic Data Attribution: Tracing the Training Origins of Interpretable LLM Units

ResearchDGX agent

arXiv:2601.21996v2 Announce Type: replace-cross Abstract: While Mechanistic Interpretability has identified interpretable circuits in LLMs, their causal origins in training data remain elusive. We int

MeCo: One-Step MeanFlow-based Corrector for Multi-Channel Speech Separation

ResearchDGX agent

arXiv:2606.09677v1 Announce Type: cross Abstract: While discriminative models for multi-channel speech separation excel in reference-based metrics, they often exhibit suboptimal human listening qualit

MedicalRec: Medical recommender system for image classification without retraining

ApplicationsDGX agent

arXiv:2606.07553v1 Announce Type: cross Abstract: The emergence of machine learning and deep learning has revolutionized the efficiency of diagnostic, therapeutic, and administrative systems in health

MedVision: Benchmarking Quantitative Medical Image Analysis

Model ReleasesDGX agent

arXiv:2511.18676v2 Announce Type: replace-cross Abstract: Current vision-language models (VLMs) in medicine are primarily designed for categorical question answering (e.g., 'Is this normal or abnormal

Meeting SLOs, Slashing Hours: Automated Enterprise LLM Optimization with OptiKIT

HardwareDGX agent

arXiv:2601.20408v2 Announce Type: replace-cross Abstract: Enterprise LLM deployment faces a critical scalability challenge: organizations must optimize models systematically to scale AI initiatives wi

Memetic Capture: A Pluralistic Policy Framework for Governing AI-Driven Cultural Disempowerment

SafetyDGX agent

arXiv:2606.07802v1 Announce Type: cross Abstract: Culture is the most insidious vector of gradual human disempowerment by AI: unlike economic or political displacement, cultural displacement attacks t

Memory Beyond Recall: A Dual-Process Cognitive Memory System for Self-Evolving LLM Agents

Model ReleasesDGX agent

arXiv:2606.09483v1 Announce Type: cross Abstract: Long-term memory for an LLM agent is more than retrieving the right passage at the right time. Current memory systems collapse belief revision, causal

MemoVAD: Resource-Efficient Video Anomaly Detection via Dynamic Semantic Memory in Edge Computing Scenarios

Local AiDGX agent

arXiv:2606.07669v1 Announce Type: cross Abstract: Deploying Video Anomaly Detection (VAD) in real-world surveillance faces a fundamental tension between the demand for high-level semantics to ensure e

MemToolAgent overview with a simple restaurant booking scenario where the agent retrieves similar memories, receives feedback on an invalid time format, and generates a reflection to update its memory

AgentsDGX agent

arXiv:2606.07909v1 Announce Type: new Abstract: Modern large language model (LLM) agents can use external tools to help users solve complex tasks. However, for problems that require learning from long

MEnvAgent: Scalable Polyglot Environment Construction for Verifiable Software Engineering

Model ReleasesDGX agent

arXiv:2601.22859v3 Announce Type: replace-cross Abstract: The evolution of Large Language Model (LLM) agents for software engineering (SWE) is constrained by the scarcity of verifiable datasets, a bot

MetaEvo: A Meta-Optimization Framework for Experience-Driven Agent Evolution

AgentsDGX agent

arXiv:2606.07603v1 Announce Type: cross Abstract: Large language models (LLMs) exhibit strong reasoning capabilities, yet most LLM-based agents are statically deployed and unable to improve through ta

Minibatch Selection via Partition Matroid Constrained Gradient Matching

Model ReleasesDGX agent

arXiv:2606.07954v1 Announce Type: cross Abstract: Training large language models (LLMs) on heterogeneous data requires selecting minibatches that balance convergence speed with coverage across domains

MIRAGE: Metadata-Integrated Repository Analysis and Guided Enhancement for MSR Datasets

SafetyDGX agent

arXiv:2606.07611v1 Announce Type: cross Abstract: This paper proposes an improved approach to the analysis of Mining Software Repositories (MSR) datasets via metadata enrichment, FAIRness assessment,

MixReasoning: Switching Modes to Think

TutorialsDGX agent

arXiv:2510.06052v2 Announce Type: replace Abstract: Reasoning models enhance performance by tackling problems in a step-by-step manner, decomposing them into sub-problems and exploring long chains of

MLingualFC: Evaluating Jailbreak Vulnerabilities in Multilingual Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.07706v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) have demonstrated strong performance across multimodal tasks, yet their safety robustness remains an open challenge. Whi

mllm-shap: A Shapley Value Explainability Platform for Text-Audio Multimodal Large Language Models

SafetyDGX agent

arXiv:2606.07531v1 Announce Type: cross Abstract: We introduce mllm-shap, an open-source Python framework designed to extend Shapley Value (SV) explainability from text-only Large Language Models to M

MM-Matryoshka: Towards Budget-Elastic Visual Document Retrieval via a 2D Multimodal Matryoshka Training Framework

ResearchDGX agent

arXiv:2606.07654v1 Announce Type: cross Abstract: Multi-vector visual document retrievers achieve strong fine-grained matching by representing each page with multiple vectors from deep Vision-Language

MMR-GRPO: Accelerating GRPO-Style Training through Diversity-Aware Reward Reweighting

Model ReleasesDGX agent

arXiv:2601.09085v2 Announce Type: replace-cross Abstract: Group Relative Policy Optimization (GRPO) has become a standard approach for training mathematical reasoning models; however, its reliance on

Mobility-Embedded POIs: Learning What A Place Is and How It Is Used from Human Movement

TutorialsDGX agent

arXiv:2601.21149v3 Announce Type: replace-cross Abstract: Recent progress in geospatial foundation models highlights the importance of learning general-purpose representations for real-world locations

Model Multiplicity for Adversarial Detection in Small Language Model Training on Edge Devices

Model ReleasesDGX agent

arXiv:2606.07857v1 Announce Type: cross Abstract: The rise of edge-based machine learning has enabled distributed adaptation of language models across mobile and IoT devices, offering privacy preserva

Model Poisoning Against Federated Model Adaptation with Chain of Bit-Flips

Local AiDGX agent

arXiv:2606.09548v1 Announce Type: cross Abstract: Federated Learning (FL) allows a set of clients to collectively train a global model without sharing local training data. Giving the responsibility of

Modeling the Diachronic Evolution of Legal Norms: An LRMoo-Based, Component-Level, Event-Centric Approach to Legal Knowledge Graphs

ApplicationsDGX agent

arXiv:2506.07853v5 Announce Type: replace Abstract: Representing the temporal evolution of legal norms is a critical challenge for automated processing. While foundational frameworks exist, they lack

Momentum for Reasoning: Dense Intrinsic Signals in Policy Optimization

SafetyDGX agent

arXiv:2606.08815v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for eliciting long-chain reasoning in large language models. Ho

More Bang for the Buck: Improving the Inference of Large Language Models at a Fixed Budget using Reset and Discard (ReD)

ResearchDGX agent

arXiv:2601.21522v2 Announce Type: replace-cross Abstract: The performance of large language models (LLMs) on verifiable tasks is usually measured by pass@k, the probability of answering a question cor

More Yap Less Meaning: Uncovering Self-Improvement Behavior in SLMs

ResearchDGX agent

arXiv:2606.08471v1 Announce Type: cross Abstract: Recently, language models have made rapid progress across various domains and applications. However, their capability for self-improvement, i.e., whet

MOSS-Video-Preview: Toward Real-Time Video Understanding via Cross-Attention

ResearchDGX agent

arXiv:2606.07639v1 Announce Type: cross Abstract: Video understanding is shifting from the offline paradigm -- taking a fully recorded video as input and producing a single answer after it ends -- tow

Multi-planar 2D-U-Net Segmentation of 3D-CT Abdominal Organs augmented by Spatial Occurrence Maps

ResearchDGX agent

arXiv:2606.07717v1 Announce Type: cross Abstract: This work proposes a lightweight 2D-U-Net-based framework for segmenting five abdominal organs in large field-of-view 3D CT scans. The method combines

Multi-Turn Evaluation of Deep Research Agents Under Process-Level Feedback

AgentsDGX agent

arXiv:2606.09748v1 Announce Type: new Abstract: Existing benchmarks for deep research agents (DRAs) assess only single-shot outputs, ignoring a key question: can DRAs improve their reports when guided

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers

SafetyDGX agent

arXiv:2601.12263v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) integrate visual and textual knowledge into unified representations that increasingly underpin modern retrieval

Multimodal Group Emotion Recognition In-the-Wild Towards a Privacy-Safe Non-Individual Approach

ApplicationsDGX agent

arXiv:2606.07585v1 Announce Type: cross Abstract: This thesis addresses group emotion recognition (GER) in-the-wild with a focus on privacy preservation. Unlike traditional emotion recognition methods

Multimodal Large Language Models as Synthetic Participants in Video-Based Studies: An Evaluation

Model ReleasesDGX agent

arXiv:2606.07541v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) have shown strong performance on objective tasks such as video understanding and reasoning. However, it remai

Muon Learns More Robust and Transferable Features than Adam

ResearchDGX agent

arXiv:2606.09658v1 Announce Type: cross Abstract: Muon has recently emerged as a state-of-the-art optimizer for pretraining Large Language Models (LLMs) and vision classifiers. Despite its efficiency

← Previous
1…145146147148149…358
Next →