AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,474 results
10 Jul 2026

ReCoLoRA: Spectrum-Aware Recursive Consolidation for Continual LLM Fine-Tuning

Model ReleasesDGX agent

arXiv:2607.07719v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning adapts a large language model to one task cheaply, but across a task sequence LoRA-style methods keep stacking low-ran

Reinforcing the Generation Order of Multimodal Masked Diffusion Models

Model ReleasesDGX agent

arXiv:2607.08056v1 Announce Type: cross Abstract: Diffusion Language Models (DLMs) have recently achieved substantial progress in natural language generation tasks. Recent research demonstrates that a

Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents

Model ReleasesDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.08716v1 Announce Type: new Abstract: In long-horizon tasks, decision-relevant state is often scattered across an expanding trajectory, while the action agent must surface it and act. As tra

RetailBench: Evaluating Long-Horizon Autonomous Decision-Making and Strategy Stability of LLM Agents in Realistic Retail Environments

Model ReleasesDGX agent

arXiv:2603.16453v3 Announce Type: replace Abstract: Large language model (LLM) agents have made rapid progress on short-horizon, well-scoped tasks, yet their ability to sustain coherent decisions in d

Rethinking LLM-as-a-Judge: Representation-as-a-Judge with Small Language Models via Semantic Capacity Asymmetry

ResearchDGX agent

arXiv:2601.22588v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are widely used as reference-free evaluators via prompting, but this 'LLM-as-a-Judge' paradigm is costly, opaque,

Retrieval of Scientific and Technological Resources for Experts and Scholars

ResearchDGX agent

arXiv:2204.06142v2 Announce Type: replace-cross Abstract: Institutions of higher learning, research institutes and other scientific research units have abundant scientific and technological resources

RhyMix: A Lightweight Adaptive Multi-Rhythm Network for Long-Term Time Series Forecasting

Local AiDGX agent

arXiv:2607.08234v1 Announce Type: cross Abstract: Real-world time series exhibit complex dynamics characterized by multiple simultaneous temporal patterns: short-term fluctuations, periodic seasonal c

Robust Weighted Triangulation of Causal Effects Under Model Uncertainty

ResearchDGX agent

arXiv:2603.01119v2 Announce Type: replace-cross Abstract: A fundamental challenge in causal inference with observational data is correct specification of a causal model. When there is model uncertaint

Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies

SafetyDGX agent

arXiv:2603.15136v2 Announce Type: replace-cross Abstract: Offline safe reinforcement learning (RL) seeks reward-maximizing policies from static datasets under strict safety constraints. Existing metho

Self-Adaptive Anomaly Detection with Reinforcement Learning and Human Feedback in Connected Vehicles

AgentsDGX agent

arXiv:2607.08373v1 Announce Type: cross Abstract: Connected vehicles are autonomous cyber-physical systems whose behavior must be continuously monitored during operation to detect deviations from norm

Self-EvolveRec: Self-Evolving Recommender Systems with LLM-based Directional Feedback

ResearchDGX agent

arXiv:2602.12612v2 Announce Type: replace-cross Abstract: Traditional methods for automating recommender system design, such as Neural Architecture Search (NAS), are often constrained by a fixed searc

SHAP-Weighted Cross-Modal Expert Fusion for Emotion and Sentiment Recognition: Evidence and Limits

ResearchDGX agent

arXiv:2607.08573v1 Announce Type: new Abstract: Multimodal emotion and sentiment recognition is commonly addressed by early fusion, which concatenates modalities before classification, or late fusion,

Shift & Drift: A Zero-Shot Benchmark for Generalizable and Robust Autonomous Driving Motion Planning

Model ReleasesDGX agent

arXiv:2607.07844v1 Announce Type: cross Abstract: While closed-loop motion planners trained on large-scale, object-level datasets, e.g., nuPlan, demonstrate strong in-distribution (ID) performance, th

SHIFT: Survival Prediction from Incomplete and Heterogeneous Genomic Data

ResearchDGX agent

arXiv:2607.07725v1 Announce Type: cross Abstract: Genomic prediction models often fail to transfer across institutions because sequencing panels differ across sites, creating structural feature missin

SimRPD: Optimizing Recruitment Proactive Dialogue Agents through Simulator-Based Data Evaluation and Selection

AgentsDGX agent

arXiv:2601.02871v3 Announce Type: replace Abstract: Task-oriented proactive dialogue agents play a pivotal role in recruitment, particularly for steering conversations towards specific business outcom

Simulator Ensembles for Trustworthy Autonomous Driving Systems Testing

AgentsDGX agent

arXiv:2503.08936v3 Announce Type: replace-cross Abstract: Scenario-based testing with driving simulators is extensively used to identify failing conditions of automated driving assistance systems (ADA

SLORR: Simple and Efficient In-Training Low-Rank Regularization

HardwareDGX agent

arXiv:2607.08754v1 Announce Type: cross Abstract: Low-rank factorization is widely used to compress neural networks, but modern models are often not naturally amenable to aggressive factorization with

SMetric: Rethink LLM Scheduling for Serving Agents with Balanced Session-centric Scheduling

AgentsDGX agent

arXiv:2607.08565v1 Announce Type: cross Abstract: LLM scheduling is critical to serving, yet it remains unclear how well existing designs fit agentic serving--with LLM requests issued by agents instea

SolarChain-Eval: A Physics-Constrained Benchmark for Trustworthy Economic Agents in Decentralized Energy Markets

Model ReleasesDGX agent

arXiv:2607.08681v1 Announce Type: new Abstract: As agentic AI systems are increasingly applied to cyber-physical environments, their evaluation requires assessment of both task performance and trustwo

Spatio-Temporal Scheduling Prediction Under Backhaul Delay for Resilient Coordinated Beamforming

SafetyDGX agent

arXiv:2607.08454v1 Announce Type: cross Abstract: Coordinated beamforming in distributed 5G networks relies on the timely exchange of inter-cell scheduling information, but backhaul latency makes this

Spectral Analysis of Dueling Q-Learning

ResearchDGX agent

arXiv:2607.08340v1 Announce Type: cross Abstract: Q-learning is a fundamental algorithm in reinforcement learning (RL) for solving discounted Markov decision processes (MDPs) when the transition kerne

SpO_2 Predictor-Guided Stage-Wise Time-Frequency Reconstruction of Low-Quality Dual-Wavelength PPG for Oxygen Saturation Estimation

ResearchDGX agent

arXiv:2607.07996v1 Announce Type: cross Abstract: Continuous oxygen saturation (SpO_2) estimation from wearable photoplethysmography (PPG) is important for long-term health monitoring, but low-quality

StateLinFormer: Stateful Training Enhancing Long-term Memory in Navigation

ResearchDGX agent

arXiv:2603.23571v2 Announce Type: replace-cross Abstract: Effective navigation intelligence relies on long-term memory to support both immediate generalization and sustained adaptation. However, exist

Structured Pruning of Large Language Models via Power Transformation and Sign-Preserving Score Aggregation with Adaptive Feature Retention

Model ReleasesDGX agent

arXiv:2607.08027v1 Announce Type: cross Abstract: This paper proposes an improved structured pruning method for large language models (LLMs) that addresses key challenges in adapting Adaptive Feature

Swapping Faces, Saving Features: A Dual-Purpose Pipeline for Pedestrian Privacy in ITS

AgentsDGX agent

arXiv:2607.08402v1 Announce Type: cross Abstract: Large-scale and diverse datasets are needed to train AI models to take real-time decisions for autonomous vehicles (AVs), an intelligent transportatio

SwinIFS: Landmark Guided Swin Transformer For Identity Preserving Face Super Resolution

Model ReleasesDGX agent

arXiv:2601.01406v2 Announce Type: replace-cross Abstract: Face super-resolution aims to recover high-quality facial images from severely degraded low-resolution inputs, but remains challenging due to

The complexities of patient-centred conversational artificial intelligence

ApplicationsDGX agent

arXiv:2607.08625v1 Announce Type: new Abstract: Consumer-facing health chatbots powered by large language models (LLMs) are increasingly used for symptom assessment. However, chatbot development and e

The Context Access Divide: Interaction-Level Architecture as a Complementary Dimension of Agentic Inequality

AgentsDGX agent

arXiv:2607.08495v1 Announce Type: cross Abstract: Sharp et al. (2025) introduce 'agentic inequality' as a framework for analyzing disparities in access to AI agents across three dimensions: availabili

The Contribution of XAI for the Safe Development and Certification of AI: An Expert-Based Analysis

SafetyDGX agent

arXiv:2408.02379v2 Announce Type: replace-cross Abstract: Developing and certifying safe - or so-called trustworthy - AI has become an increasingly salient issue, especially in light of upcoming regul

The Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMs

ApplicationsDGX agent

arXiv:2607.08734v1 Announce Type: new Abstract: Post-training quantization is widely used to deploy large language models in resource-constrained settings, yet its evaluation relies almost exclusively

The Phasor Transformer: Resolving Attention Bottlenecks on the Unit Circle

Model ReleasesDGX agent

arXiv:2603.17433v2 Announce Type: replace-cross Abstract: Transformer models have redefined sequence learning, yet dot-product self-attention introduces a quadratic token-mixing bottleneck for long-co

Time-to-Collision Based Dynamic Obstacle Avoidance Using Pretrained Vision Models for Robots in Unstructured Environments

AgentsDGX agent

arXiv:2607.07885v1 Announce Type: cross Abstract: Dynamic obstacle avoidance in unstructured outdoor environments remains a critical challenge for autonomous mobile robots, particularly when large-sca

TMI: Text-to-Image Meets Image-to-Image for Complementary Data Synthesis to Boost Long-Tailed Instance Segmentation

Model ReleasesDGX agent

arXiv:2607.08201v1 Announce Type: cross Abstract: Large-vocabulary instance segmentation is constrained by long-tailed category distributions and fine-grained inter-class ambiguity. While data synthes

TNODEV: Toolbox for Neural ODE Verification

SafetyDGX agent

arXiv:2606.16567v2 Announce Type: replace Abstract: Neural ordinary differential equations (neural ODE) gained attention in safety critical settings such as continuous-time controllers for cyber-physi

ToDMA: Large Model-Driven Massive Token Communications for Semantic Multiple Access

ResearchDGX agent

arXiv:2505.10946v3 Announce Type: replace-cross Abstract: Token communications (TokenCom) is an emerging generative semantic communication paradigm, where tokens serve as compact representation units

TOPO-Bench: An Open-Source Topological Mapping Evaluation Framework with Quantifiable Perceptual Aliasing

Model ReleasesDGX agent

arXiv:2510.04100v2 Announce Type: replace-cross Abstract: Topological mapping offers a compact and robust representation for navigation, but progress in the field is hindered by the lack of standardiz

Towards Efficient Large Language Model Serving: A Survey on System-Aware KV Cache Optimization

ResearchDGX agent

arXiv:2607.08057v1 Announce Type: cross Abstract: Despite the rapid advancements of large language models (LLMs), LLM serving systems remain memory-intensive and costly. The key-value (KV) cache, whic

Towards Isolated Interventions via Almost Orthogonal Features in Language Models

Local AiDGX agent

arXiv:2602.04718v2 Announce Type: replace-cross Abstract: A central premise in mechanistic interpretability is that meaningful concepts in language models are represented by linear features in activat

Towards Mechanistically Understanding Why Memorized Knowledge Fails to Generalize in Large Language Model Finetuning

ResearchDGX agent

arXiv:2607.08393v1 Announce Type: new Abstract: Fine-tuning LLMs to inject new knowledge faces a critical challenge: LLMs can quickly memorize new facts, yet fail to use them for downstream reasoning

Towards Precision Therapy in Hepatocellular Carcinoma: A Clinical-Reasoning LLM for Risk Stratification and Treatment Guidance

Model ReleasesDGX agent

arXiv:2607.08602v1 Announce Type: new Abstract: Hepatocellular carcinoma (HCC) is a common malignancy and a leading cause of cancer-related mortality. Current guidelines and staging systems provide co

Towards the Explainability of Temporal Graph Networks via Memory Backtracking and Topological Attribution

ApplicationsDGX agent

arXiv:2607.07716v1 Announce Type: cross Abstract: Temporal graphs are ubiquitous in real-world applications and Temporal Graph Networks (TGNs) have achieved superior predictive accuracy. Understanding

TRACE: A Two-Channel Robust Attribution Watermark via Complementary Embeddings for LLM-Agent Trajectories

Local AiDGX agent

arXiv:2607.08400v1 Announce Type: cross Abstract: LLM agents reach users through resellers, who may rebrand a developer's agent or substitute a cheaper model. When provenance is disputed, attribution

Track2Map: Online Deformable SLAM with Motion-Aware Pose Optimization in Robotic Surgery

ResearchDGX agent

arXiv:2607.08408v1 Announce Type: cross Abstract: Gaussian splatting is the current state-of-the-art for dense, deformable 3D anatomy reconstruction in robot-assisted minimally invasive surgery (RAMIS

Training and Evaluating Diffusion Policies with Long Context Lengths

Model ReleasesDGX agent

arXiv:2606.16447v2 Announce Type: replace-cross Abstract: Imitation learning has enabled highly-dexterous robotic manipulation from RGB observations. Policies trained with these methods, however, typi

Two Axes of LLM Abstention: Answer Correctness and Question Answerability

SafetyDGX agent

arXiv:2607.08456v1 Announce Type: cross Abstract: A model should refuse two different things: answers it would get wrong, and questions it should not answer at all, such as unanswerable ones or ones r

TypeProbe: Recovering Type Representations from Hidden States of Pre-trained Code Models

ResearchDGX agent

arXiv:2607.08339v1 Announce Type: cross Abstract: State-of-the-art code models achieve impressive performance, yet the extent to which they internally encode type information remains poorly understood

UltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic Editing

SafetyDGX agent

arXiv:2607.08646v1 Announce Type: cross Abstract: As available training data approaches its physical limit, gains from Scaling Laws have begun to diminish. Consequently, improving Large Language Model

Understanding Axes of Difficulty For Long Context Tasks Via PredicateLongBench

Model ReleasesDGX agent

arXiv:2607.08284v1 Announce Type: new Abstract: Large language models (LLMs) have demonstrated rapidly improving long-context capabilities, prompting a wave of benchmarks designed to evaluate them. Ho

Using AI-based Learning Assistants in Higher Education: A Large-Scale Descriptive Analysis

ApplicationsDGX agent

arXiv:2607.08748v1 Announce Type: new Abstract: In this study, we present a large-scale descriptive analysis of the use of an AI-based learning assistant (Syntea) in higher education. Based on objecti

Validity of LLMs as data annotators: AMALIA on authority

Model ReleasesDGX agent

arXiv:2607.08731v1 Announce Type: cross Abstract: A national language model offers a linguistic community its own instrument for measuring what its citizens say and value. Portugal's AMALIA, a publicl

VectorizationLLM: Smart Vectorization Based AI Assistant

TutorialsDGX agent

arXiv:2607.07846v1 Announce Type: new Abstract: VectorizationLLM is a specialized Large Language Model based on Google open-weight LLMs. The model is designed to assist students to learn smart vectori

VEGAS: Human-Aligned Video Caption Evaluation via Gaze

ResearchDGX agent

arXiv:2607.08489v1 Announce Type: cross Abstract: Vision-language models excel at video captioning, yet typically generate descriptions that fail to capture individual viewers' attention. We propose V

VocaDet: Sample-Driven Open-Vocabulary Object Detection and Segmentation via Visual Tokenization and Vector Database Retrieval

ResearchDGX agent

arXiv:2607.08541v1 Announce Type: cross Abstract: Open-vocabulary object detection and segmentation aim to recognize arbitrary objects beyond predefined categories. Although recent vision-language and

WCog-VLA: A Dual-Level World-Cognitive Vision-Language-Action Model for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2607.08375v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have advanced end-to-end autonomous driving. However, existing methods either lack comprehensive world cognition o

WebSwarm: Recursive Multi-Agent Orchestration for Deep-and-Wide Web Search

Local AiDGX agent

arXiv:2607.08662v1 Announce Type: cross Abstract: Large language model (LLM)-based web search agents are transforming information seeking from simple factoid question answering into complex, deep-and-

What LLM Forecasters Know but Don't Say: Probing Internal Representations for Calibration and Faithfulness

ResearchDGX agent

arXiv:2607.08046v1 Announce Type: cross Abstract: Large language models fine-tuned for forecasting can be accurate yet poorly calibrated, and their chain-of-thought (CoT) reasoning may not faithfully

When Implausible Tokens Get Reinforced: Tail-Aware Credit Calibration for LLM Reinforcement Learning

ResearchDGX agent

arXiv:2607.07976v1 Announce Type: cross Abstract: Reinforcement learning (RL) has achieved remarkable success in enhancing the reasoning capabilities of large language models (LLMs). However, widely u

When LLMs Agree, Are They Right? Auditing Self-Consistency and Cross-Model Agreement as Confidence Signals

Model ReleasesDGX agent

arXiv:2607.08065v1 Announce Type: new Abstract: LLM-as-judge (Zheng et al., 2023) is increasingly the default for evaluating AI systems in enterprise pipelines, often scaled to ensembles (Verga et al.

When Structured Sparse Autoencoders Learn Consistent Concepts Across Modalities

SafetyDGX agent

arXiv:2607.08605v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have emerged as a promising technique for mechanistic interpretability by learning a set of sparse latent features in large

When Synthetic Speech Is All You Have: Better Call GRPO

SafetyDGX agent

arXiv:2607.08409v1 Announce Type: cross Abstract: LLM-based ASR adapted to regulated domains such as banking is bottlenecked by privacy: real speech is costly and legally constrained to collect, makin

← Previous
1…7576777879…358
Next →