AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,136
  • Agents7,313
  • Applications5,230
  • Concepts5
  • Hardware1,765
  • Industry6,107
  • Local Ai4,758
  • Model Releases22,770
  • Research19,333
  • Safety12,890
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlog
85,136Total entries
1Added by human
85,135Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ai”

GridTimelineEvolution
21,688 results
Model Releases

Less is Enough: Synthesizing Diverse Data in LLM Feature Space with Sparse Autoencoders

DGX agent

arXiv:2602.10388v3 Announce Type: replace-cross Abstract: The diversity of post-training data is critical for effective downstream performance in large language models (LLMs). Many existing approaches

model-releasesarxiv-cs-ai
29 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Less Is More: Elevating RAG via Performance-Driven Context Compression

DGX agent

arXiv:2508.19282v4 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has emerged as a promising paradigm for improving the timeliness of knowledge updates and the factual acc

safetyarxiv-cs-ai
29 May 2026
Research

LFQ: Logit-aware Final-block Quantization for Boosting the Generation Quality of Low-Bit Quantized LLMs

DGX agent

arXiv:2605.29756v1 Announce Type: new Abstract: As large language models continue to scale, low-bit weight-only post-training quantization (PTQ) offers a practical solution to their memory-efficient d

researcharxiv-cs-ai
29 May 2026
Model Releases

LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning

DGX agent

arXiv:2605.29649v1 Announce Type: new Abstract: Heuristic search is the dominant paradigm in symbolic AI planning, and the strongest heuristics are the result of decades of work by planning researcher

model-releasesarxiv-cs-ai
29 May 2026
Research

LLMSurgeon: Diagnosing Data Mixture of Large Language Models

DGX agent

arXiv:2605.30348v1 Announce Type: cross Abstract: The pretraining data mixture of Large Language Models (LLMs) constitutes their 'digital DNA', shaping model behaviors, capabilities, and failure modes

researcharxiv-cs-ai
29 May 2026
Safety

LLUMI: Improving LLM Writing Assistance for Mental Health Support with Online Community Feedback

DGX agent

arXiv:2605.30273v1 Announce Type: cross Abstract: Large language models (LLMs) show promise in generating supportive responses for mental health queries, but improving their usefulness, empathy, and s

safetyarxiv-cs-ai
29 May 2026
Local Ai

Locally Coherent, Globally Incoherent: Bounding Compositional Incoherence in Multi-Component LLM Agents

DGX agent

arXiv:2605.30335v1 Announce Type: new Abstract: Multi-component LLM agents assemble probabilistic claims from components that each see only part of a joint problem; the composition can violate basic p

local-aiarxiv-cs-ai
29 May 2026
Model Releases

LoCoT2V-Bench: Benchmarking Long-Form and Complex Text-to-Video Generation

DGX agent

arXiv:2510.26412v3 Announce Type: replace-cross Abstract: Recent advances in text-to-video generation have achieved impressive performance on short clips, yet evaluating long-form generation under com

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LogDx-CI: Benchmarking Log Reduction Tools for LLM Root-Cause Diagnosis

DGX agent

arXiv:2605.28876v1 Announce Type: cross Abstract: CI failure logs are large (median 5k lines, max 200k in this corpus) and noisy. Coding agents that try to debug them depend on an upstream tool to red

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Loong: A Human-Like Long Document Translation Agent with Observe-and-Act Adaptive Context Selection

DGX agent

arXiv:2605.30274v1 Announce Type: cross Abstract: Document-level translation remains one of the most challenging tasks for large language models, which are constrained by limited context windows that

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LoopFM: Learning frOm HistOrical RePresentations of Foundation Model for Recommendation

DGX agent

arXiv:2605.29280v1 Announce Type: cross Abstract: Knowledge distillation (KD) transfers a single scalar prediction from a large foundation model (FM) to compact vertical models (VMs), suffering from d

model-releasesarxiv-cs-ai
29 May 2026
Research

LoRe: Adaptive Interaction-Evaluation Routing with Per-Step Interaction Budgets for Iterative Graph Solvers

DGX agent

arXiv:2605.29005v1 Announce Type: cross Abstract: Diffusion-based neural solvers for combinatorial optimization repeatedly re-evaluate dense edge/factor interactions, making inference expensive in wal

researcharxiv-cs-ai
29 May 2026
Applications

LsrIF: Enhancing Logic-Structured Instruction Following of Large Language Models

DGX agent

arXiv:2601.06431v3 Announce Type: replace Abstract: Instruction following is critical for large language models, yet real-world instructions often involve multiple constraints with logical structures,

applicationsarxiv-cs-ai
29 May 2026
Tutorials

Make LLM Learn to Synthesize from Streaming Experiences through Feedback

DGX agent

arXiv:2605.29940v1 Announce Type: new Abstract: Large language models (LLMs) have been widely adopted for synthetic data generation, significantly reducing annotation costs. However, most existing stu

tutorialsarxiv-cs-ai
29 May 2026
Safety

Masked Diffusion Modeling for Anomaly Detection

DGX agent

arXiv:2605.30046v1 Announce Type: cross Abstract: Anomaly detection aims to identify samples that deviate from the nominal data distribution and is central to many safety-critical applications. Howeve

safetyarxiv-cs-ai
29 May 2026
Model Releases

MATNet: Multi-Level Fusion Transformer-Based Model for Day-Ahead PV Generation Forecasting

DGX agent

arXiv:2306.10356v3 Announce Type: replace-cross Abstract: Accurate forecasting of renewable generation is crucial to facilitate the integration of Renewable Energy Sources into the power system. Focus

model-releasesarxiv-cs-ai
29 May 2026
Agents

mcp-proto-okn: Natural-language access to open scientific knowledge graphs through the Model Context Protocol

DGX agent

arXiv:2605.30283v1 Announce Type: new Abstract: MCP Server Proto-OKN (mcp-proto-okn) is a Python-based Model Context Protocol server that enables AI assistants to discover, inspect, query and integrat

agentsarxiv-cs-ai
29 May 2026
Applications

Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening

DGX agent

arXiv:2605.28999v1 Announce Type: cross Abstract: LLMs are vulnerable to prompt injection attacks. However, this vulnerability has been primarily demonstrated conceptually in academic studies or throu

applicationsarxiv-cs-ai
29 May 2026
Research

Mechanism Shift During Post-training from Autoregressive to Masked Diffusion Language Models

DGX agent

arXiv:2601.14758v4 Announce Type: replace-cross Abstract: Post-training pretrained autoregressive models (ARMs) into masked diffusion models (MDMs) has emerged as a cost-effective way to overcome the

researcharxiv-cs-ai
29 May 2026
Model Releases

Mechanistic origins of catastrophic forgetting: why RL preserves circuits better than SFT?

DGX agent

arXiv:2605.28860v1 Announce Type: cross Abstract: Fine-tuning large language models (LLMs) frequently induces catastrophic forgetting of prior capabilities. Recent work has shown that reinforcement le

model-releasesarxiv-cs-ai
29 May 2026
Research

MedCase-Structured: A Text-to-FHIR Dataset for Benchmarking Diagnostic Reasoning in Clinically Realistic EHR Settings

DGX agent

arXiv:2605.30295v1 Announce Type: cross Abstract: Large language models (LLMs) show promise for clinical reasoning and decision support, but evaluation in realistic, electronic health record-congruent

researcharxiv-cs-ai
29 May 2026
Local Ai

MediHive: A Decentralized Agent Collective for Medical Reasoning

DGX agent

arXiv:2603.27150v2 Announce Type: replace Abstract: Large language models (LLMs) have revolutionized medical reasoning tasks, yet single-agent systems often falter on complex, interdisciplinary proble

local-aiarxiv-cs-ai
29 May 2026
Agents

MemCollab: Cross-Model Memory Collaboration via Contrastive Trajectory Distillation

DGX agent

arXiv:2603.23234v2 Announce Type: replace Abstract: LLM agents increasingly rely on memory mechanisms to reuse knowledge from past problem-solving experiences. However, existing methods typically cons

agentsarxiv-cs-ai
29 May 2026
Tutorials

MEMENTO: Leveraging Web as a Learning Signal for Low-Data Domains

DGX agent

arXiv:2605.29795v1 Announce Type: new Abstract: Real-world tasks often lack large labeled datasets, motivating extensive work on learning in low-data regimes. However, existing approaches such as few-

tutorialsarxiv-cs-ai
29 May 2026
Research

MemoSight: Unifying Context Compression and Multi Token Prediction for Reasoning Acceleration

DGX agent

arXiv:2604.14889v2 Announce Type: replace Abstract: While chain-of-thought (CoT) reasoning enables LLMs to solve challenging reasoning tasks, the linear growth of the KV cache leads to substantial mem

researcharxiv-cs-ai
29 May 2026
Model Releases

MENTOR: Efficient Multimodal-Conditioned Tuning for Autoregressive Vision Generation Models

DGX agent

arXiv:2507.09574v3 Announce Type: replace-cross Abstract: Recent text-to-image models produce high-quality results but still struggle with precise visual control, balancing multimodal inputs, and requ

model-releasesarxiv-cs-ai
29 May 2026
Local Ai

Meta-Cognitive Memory Policy Optimization for Long-Horizon LLM Agents

DGX agent

arXiv:2605.30159v1 Announce Type: new Abstract: Memory-augmented LLM agents tackle complex long-horizon tasks by recursively summarizing interaction trajectories into compact memory. However, existing

local-aiarxiv-cs-ai
29 May 2026
Research

Meta-Programming for Linear-time Temporal Answer Set Programming

DGX agent

arXiv:2605.29965v1 Announce Type: new Abstract: The development of temporal extensions of Answer Set Programming (ASP) has led to the emergence of non-monotonic linear-time (TEL), dynamic (DEL), and m

researcharxiv-cs-ai
29 May 2026
Research

MiAD: Mirage Atom Diffusion for De Novo Crystal Generation

DGX agent

arXiv:2511.14426v2 Announce Type: replace-cross Abstract: In recent years, diffusion-based models have demonstrated exceptional performance in searching for simultaneously stable, unique, and novel (S

researcharxiv-cs-ai
29 May 2026
Research

Micro-Macro Retrieval: Reducing Long-Form Hallucination in Large Language Models

DGX agent

arXiv:2605.28828v1 Announce Type: cross Abstract: Large Language Models (LLMs) achieve impressive performance across many tasks but remain prone to hallucination, especially in long-form generation wh

researcharxiv-cs-ai
29 May 2026
Research

Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete Diffusion

DGX agent

arXiv:2605.29591v1 Announce Type: new Abstract: Modeling the interplay between external stimuli and internal neural representations is a pivotal research area for Brain-Computer Interfaces (BCIs). A m

researcharxiv-cs-ai
29 May 2026
Model Releases

Mind Your Tone: Does Tone Alter LLM Performance?

DGX agent

arXiv:2605.29027v1 Announce Type: new Abstract: The use of Large Language Models (LLMs) is proliferating, yet their performance is observed to vary based on prompting styles and tones. In this study,

model-releasesarxiv-cs-ai
29 May 2026
Agents

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs

DGX agent

arXiv:2605.29512v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as interactive agents, yet their capacity for social and strategic reasoning over extended intera

agentsarxiv-cs-ai
29 May 2026
Research

MIRA: Mid-training Rubric Anchoring for Source-Aware Data Selection

DGX agent

arXiv:2605.30288v1 Announce Type: new Abstract: Mid-training has become an important stage in modern LLM development, using large-scale curated mixtures to strengthen capabilities before final post-tr

researcharxiv-cs-ai
29 May 2026
Model Releases

MiraBench: Evaluating Action-Conditioned Reliability in Robotic World Models

DGX agent

arXiv:2605.29360v1 Announce Type: new Abstract: Action-conditioned world models are increasingly used as scalable simulators for robot learning, yet current evaluations provide limited evidence that t

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Mitigating Hallucination in Vision-Language Models through Barrier-Regulated Adaptive Closed-form Steering

DGX agent

arXiv:2605.29881v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) often hallucinate objects that are not present in the input image, largely because visual grounding weakens as de

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Mitigating Stethoscope-Induced Shortcuts in Respiratory Sound Classification under Federated Domain Generalization with Causality-Inspired Interventions

DGX agent

arXiv:2605.29862v1 Announce Type: cross Abstract: AI-driven respiratory sound classification (RSC) is promising for automated pulmonary disease detection, yet multi-site deployment is hindered by inte

model-releasesarxiv-cs-ai
29 May 2026
Safety

Model Fusion via Retrofitting

DGX agent

arXiv:2507.00037v2 Announce Type: replace-cross Abstract: Model fusion seeks to combine independently trained neural networks into a single model without retraining, but is complicated by representati

safetyarxiv-cs-ai
29 May 2026
Safety

Modeling Hierarchical Thinking in Large Reasoning Models

DGX agent

arXiv:2510.22437v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) solve complex tasks by generating long Chain-of-Thought (CoT) sequences; however, the emergent dynamics governing reas

safetyarxiv-cs-ai
29 May 2026
Safety

Modularizing Educational LLM-Agency for Fostering Responsible Learning Assistance

DGX agent

arXiv:2605.30187v1 Announce Type: new Abstract: The widespread adoption of AI chatbots in education will drastically change learning, making responsible deployment a critical concern. While large lang

safetyarxiv-cs-ai
29 May 2026
Research

Moment-KV: Momentum-Based Decode-Time KV Cache Compression for Long Generation

DGX agent

arXiv:2605.29873v1 Announce Type: new Abstract: Key-Value (KV) cache remains a major bottleneck for deploying Large Language Models (LLMs) in long-generation tasks. Prior work often applies uniform co

researcharxiv-cs-ai
29 May 2026
Applications

MOO: A Multi-view Oriented Observations Dataset for Viewpoint Analysis in Cattle Re-Identification

DGX agent

arXiv:2603.04314v2 Announce Type: replace-cross Abstract: Animal re-identification (ReID) faces critical challenges due to viewpoint variations, particularly in Aerial-Ground (AG-ReID) settings where

applicationsarxiv-cs-ai
29 May 2026
Agents

MOOSE-Copilot: A Web-Based Interactive Assistant for Unified Exploratory and Fine-Grained Scientific Hypothesis Discovery

DGX agent

arXiv:2605.29475v1 Announce Type: cross Abstract: Large language models (LLMs) show remarkable potential in scientific hypothesis discovery. However, existing approaches face two critical limitations:

agentsarxiv-cs-ai
29 May 2026
Model Releases

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing

DGX agent

arXiv:2605.22100v2 Announce Type: replace Abstract: Document parsing converts visually rich documents into machine-readable structured representations, forming a crucial foundation for information sys

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Multi-Legal-Bench: Evaluating LLMs on Legal Reasoning Across Jurisdictions, Languages, and Legal Traditions

DGX agent

arXiv:2605.29738v1 Announce Type: cross Abstract: Legal NLP benchmarks overwhelmingly evaluate a single language or aggregate tasks that differ fundamentally across jurisdictions, making cross-lingual

model-releasesarxiv-cs-ai
29 May 2026
Applications

Multi-Level Barriers to Generative AI Adoption Across Disciplines and Professional Roles in Higher Education

DGX agent

arXiv:2603.27052v2 Announce Type: replace-cross Abstract: Generative Artificial Intelligence (GenAI) is rapidly reshaping higher education, yet barriers to its adoption across different disciplines an

applicationsarxiv-cs-ai
29 May 2026
Safety

Multi-Resolution End-to-End Deep Neural Network for Optimizing Latency-Accuracy Tradeoff in Autonomous Driving

DGX agent

arXiv:2605.29138v1 Announce Type: cross Abstract: Latency-accuracy tradeoffs are fundamental in real-time applications of deep neural networks (DNNs) for cyber-physical systems. In autonomous driving,

safetyarxiv-cs-ai
29 May 2026
Model Releases

MuPHI: Learning Implicit Multimodal Harm Reasoning via Semantically Grounded Reward Optimization

DGX agent

arXiv:2605.29951v1 Announce Type: new Abstract: Understanding how harm emerges from interaction between otherwise benign image-text pairs requires intent-aware cross-modal reasoning beyond surface-lev

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…244245246247248…452
Next →