AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries87,042
  • Agents7,457
  • Applications5,325
  • Concepts5
  • Hardware1,802
  • Industry6,143
  • Local Ai4,863
  • Model Releases23,393
  • Research19,837
  • Safety13,180
  • Syntheses17
  • Tools1,671
  • Tutorials3,349

Source
HumanDGX agent

Content type
87,042Total entries
1Added by human
87,041Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Safety

Towards automated data analysis: A guided framework for LLM-based risk estimation

DGX agent

arXiv:2603.04631v2 Announce Type: replace Abstract: Large Language Models (LLMs) are increasingly integrated into critical decision-making pipelines, a trend that raises the demand for robust and auto

safetyarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Tree of Thoughts as a Classical Heuristic Search Problem: Formal Foundations and Design Patterns

DGX agent

arXiv:2605.28566v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable reasoning capabilities, yet their standard generation process -- auto-regressive token predict

researcharxiv-cs-ai
28 May 2026
Model Releases

Trinity: Unifying Class-Agnostic Terrain and Semantic Segmentation for Unstructured Outdoor Environments by Leveraging Synthetic Data

DGX agent

arXiv:2605.27644v1 Announce Type: cross Abstract: Terrain understanding is fundamental for mobile robots operating in unstructured outdoor environments. Existing vision-based traversability estimation

model-releasesarxiv-cs-ai
28 May 2026
Research

UNIQUE: Universal Top-k Sparse Attention for Training-free Inference and Sparsity-aware Training

DGX agent

arXiv:2605.27740v1 Announce Type: new Abstract: Long-context inference in large language models (LLMs) is bottlenecked by the linear growth of the self-attention key-value (KV) cache. Top-k sparse att

researcharxiv-cs-cl
28 May 2026
Safety

Verified Misguidance: Measuring Structural Citation Failures in Search-Augmented LLMs

DGX agent

arXiv:2605.28565v1 Announce Type: cross Abstract: Users of search-augmented LLMs rely on citations as evidence that responses are grounded in real sources, and rarely verify the cited pages themselves

safetyarxiv-cs-ai
28 May 2026
Model Releases

VeriTrip: A Verifiable Benchmark for Travel Planning Agents over Unstructured Web Corpora

DGX agent

arXiv:2605.28683v1 Announce Type: new Abstract: Existing benchmarks have laid the foundation for travel planning agents by establishing API-centric paradigms. However, as the capabilities of Autonomou

model-releasesarxiv-cs-ai
28 May 2026
Research

When prompt perturbations break your A/B test: A valid statistical test for generative surveying

DGX agent

arXiv:2605.27463v1 Announce Type: cross Abstract: Generative surveying -- where collections of LLM-based personas provide feedback on messages -- has emerged as a cheap and scalable alternative to tra

researcharxiv-cs-ai
28 May 2026
Safety

Where Rollouts Begin: Low-Load, High-Leverage First-Token Diversification for RLVR

DGX agent

arXiv:2605.28295v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) trains reasoning models without labeled trajectories, relying on grouped rollouts to expose the po

safetyarxiv-cs-ai
28 May 2026
Research

Which Heads Matter for Reasoning? RL-Guided KV Cache Compression

DGX agent

arXiv:2510.08525v3 Announce Type: replace Abstract: Reasoning large language models exhibit complex reasoning behaviors via extended chain-of-thought generation that are highly fragile to information

researcharxiv-cs-cl
28 May 2026
Model Releases

You Live More Than Once: Towards Hierarchical Skill Meta-Evolving

DGX agent

arXiv:2605.28390v1 Announce Type: new Abstract: Test-time skill evolving is regarded as a new paradigm for enhancing deployed agentic systems. Existing works mainly focus on hard-coded skill evolving

model-releasesarxiv-cs-ai
28 May 2026
Research

Zipping the Thought: When and How Compressed Reasoning Data Works in LLM Post-Training

DGX agent

arXiv:2605.28008v1 Announce Type: new Abstract: Large language models (LLMs) can now solve complex problems through long chain-of-thought (CoT) reasoning, but the trade-off between performance and tok

researcharxiv-cs-ai
28 May 2026
Research

A multifractal-based masked auto-encoder: an application to medical images

DGX agent

arXiv:2605.26287v1 Announce Type: new Abstract: Masked autoencoders (MAE) have shown great promise in medical image classification. However, the random masking strategy employed by traditional MAEs ma

researcharxiv-cs-cv
27 May 2026
Model Releases

AGORA: Adapter-Grounded Observation-Action Retention for Inference-Free Prompt Compression in LLM Agents

DGX agent

arXiv:2605.26596v1 Announce Type: new Abstract: The token-level extractive compressors widely used for general LM context are structurally inappropriate for LLM agents: across 17 (env, backbone, metho

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

AI evaluation may bias perceptions: The importance of context in interpreting academic writing

DGX agent

arXiv:2605.26662v1 Announce Type: cross Abstract: This paper examines how estimates of AI use in scientific writing can be biased when evaluation methods ignore contextual differences across countries

model-releasesarxiv-cs-ai
27 May 2026
Safety

Alignment Tampering: How Reinforcement Learning from Human Feedback Is Exploited to Optimize Misaligned Biases

DGX agent

arXiv:2605.27355v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is the standard method to align Large Language Models (LLMs) with human preferences. In this work, we

safetyarxiv-cs-ai
27 May 2026
Safety

Annotator Positionality as Signal: Psychometric Weighting for Anti-Autistic Ableism Detection

DGX agent

arXiv:2605.26397v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used in decision-making tasks where they can amplify or suppress perspectives, raising concerns in high-

safetyarxiv-cs-ai
27 May 2026
Model Releases

AutoDFT: A Closed-Loop Multi-Agent Framework for Autonomous DFT Calculations

DGX agent

arXiv:2605.26179v1 Announce Type: cross Abstract: Density functional theory (DFT) serves as the basis for computational discovery in materials science and chemistry, yet each calculation demands exten

model-releasesarxiv-cs-ai
27 May 2026
Safety

BAIT: Boundary-Guided Disclosure Escalation via Self-Conditioned Reasoning

DGX agent

arXiv:2605.27110v1 Announce Type: cross Abstract: In this work, we propose BAIT (Boundary-Aware Iterative Trap), a three-step jailbreak framework that approaches malicious goals through internal discl

safetyarxiv-cs-cl
27 May 2026
Model Releases

BEAT: Rhythm-Elastic Alignment for Agentic Music-guided Movie Trailer Generation

DGX agent

arXiv:2605.27067v1 Announce Type: new Abstract: Automatic movie trailer generation must select shots from a full-length film and synchronize them with background music. Existing methods either relegat

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Bridging Classification and Reconstruction: Cooperative Time Series Anomaly Detection

DGX agent

arXiv:2605.26193v1 Announce Type: cross Abstract: Time series anomaly detection (TSAD) has long been a hot research topic in data mining due to its various applications. Recent studies challenge the e

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Cesarean Scar Defect Segmentation in Transvaginal Ultrasound Images: a Dataset and Benchmark

DGX agent

arXiv:2605.26774v1 Announce Type: new Abstract: Cesarean Scar Defect (CSD) is one of the most prevalent complications following cesarean delivery. Transvaginal ultrasonography is widely used for prima

model-releasesarxiv-cs-cv
27 May 2026
Safety

CFG-OEC: Classifier Free Guidance with Orthogonal Error Correction

DGX agent

arXiv:2511.14075v2 Announce Type: replace-cross Abstract: Classifier free guidance is a standard method for conditional sampling in diffusion models, but its sampling rule is not aligned with the obje

safetyarxiv-cs-ai
27 May 2026
Tutorials

Chaos-SSL: An Attention-Based Self-Supervised Learning Framework with Chaotic Transformation for Medical Image Classification

DGX agent

arXiv:2605.27146v1 Announce Type: new Abstract: Self-Supervised Learning (SSL) has emerged as a powerful paradigm to mitigate the reliance on large, annotated datasets, a common bottleneck in medical

tutorialsarxiv-cs-cv
27 May 2026
Model Releases

CIRCLED: A Multi-turn CIR Dataset with Consistent Dialogues across Domains

DGX agent

arXiv:2605.26734v1 Announce Type: new Abstract: Existing Multi-Turn Composed Image Retrieval (MTCIR) datasets lack dialogue-history consistency and are restricted to the fashion domain. To address the

model-releasesarxiv-cs-cv
27 May 2026
Model Releases

Clinically-Grounded Counterfactual Reasoning for Medical Video Diagnosis

DGX agent

arXiv:2605.26483v1 Announce Type: new Abstract: Medical video diagnosis involves inferring clinical decisions from dynamic tissue responses throughout examination processes. Existing methods rely on a

model-releasesarxiv-cs-cv
27 May 2026
Research

Cordyceps: Covert Control Attacks on LLMs via Data Poisoning

DGX agent

arXiv:2605.26595v1 Announce Type: cross Abstract: Large language models (LLMs) are often fine-tuned on uncurated text datasets that adversaries can poison. Existing poisoning attacks primarily rely on

researcharxiv-cs-ai
27 May 2026
Safety

Counterfactual Credit Policy Optimization for Multi-Agent Collaboration

DGX agent

arXiv:2603.21563v2 Announce Type: replace Abstract: Collaborative multi-agent large language models (LLMs) can solve complex reasoning tasks by decomposing roles, but reinforcement learning for such s

safetyarxiv-cs-ai
27 May 2026
Model Releases

Deep-layer limit and stability analysis of the basic forward-backward-splitting induced network (II): learning problems

DGX agent

arXiv:2605.27133v1 Announce Type: cross Abstract: Deep unfolding neural networks derived from iterative optimization schemes and numerical ordinary/partial differential equations (ODEs/PDEs) have attr

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

DelowlightSplat: Feed-Forward Gaussian Splatting for Lowlight 3D Scene Reconstruction

DGX agent

arXiv:2605.26629v1 Announce Type: new Abstract: Novel-view synthesis and 3D reconstruction from sparse posed images are central to robotics and AR/VR. Yet, feed-forward 3D Gaussian reconstruction fail

model-releasesarxiv-cs-cv
27 May 2026
Tutorials

Dissecting Multimodal In-Context Learning: Modality Asymmetries and Circuit Dynamics in modern Transformers

DGX agent

arXiv:2601.20796v2 Announce Type: replace Abstract: Transformer-based multimodal large language models often exhibit in-context learning (ICL) abilities. Motivated by this phenomenon, we ask: how do t

tutorialsarxiv-cs-cl
27 May 2026
Model Releases

Distribution-Aware Conformal Prediction: A Framework for generating efficient prediction intervals for time series

DGX agent

arXiv:2605.26569v1 Announce Type: new Abstract: We present Distribution-aware Conformal Prediction (DCP), a unified framework integrating probabilistic predictors like Monte Carlo dropout, deep ensemb

model-releasesarxiv-cs-lg
27 May 2026
Model Releases

E3: Issue-Level Backtesting for Automated Research Critique

DGX agent

arXiv:2605.27072v1 Announce Type: cross Abstract: We present E3, an automated review assistant that augments reviewers and engineering teams by identifying decision-relevant technical concerns in rese

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

EHRSummarizer: A Privacy-Aware, FHIR-Native Reference Architecture for Source-Grounded EHR Summarization

DGX agent

arXiv:2601.01668v2 Announce Type: replace-cross Abstract: Clinicians routinely navigate fragmented electronic health record (EHR) interfaces to assemble a coherent picture of a patient's problems, med

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents

DGX agent

arXiv:2605.27240v1 Announce Type: new Abstract: Memory-augmented language agents are increasingly deployed in affective applications such as emotional support, where understanding and responding to us

model-releasesarxiv-cs-cl
27 May 2026
Agents

EvoEmo: Towards Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation

DGX agent

arXiv:2509.04310v4 Announce Type: replace Abstract: Recent research on Chain-of-Thought (CoT) reasoning in Large Language Models (LLMs) has demonstrated that agents can engage in extit{complex}, extit

agentsarxiv-cs-ai
27 May 2026
Model Releases

Exploiting Local Dynamics Regularity for Reusable Skills in Offline Hierarchical RL

DGX agent

arXiv:2605.26371v1 Announce Type: new Abstract: Hierarchical Reinforcement Learning (HRL) promises to solve long-horizon Reinforcement Learning (RL) tasks more efficiently than non-hierarchical counte

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

FAB-Bench: A Framework for Adaptive RAG Benchmarking in Semiconductor Manufacturing

DGX agent

arXiv:2605.26476v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) has become critical for knowledge-intensive applications, yet evaluating its performance in vertical domains remain

model-releasesarxiv-cs-cl
27 May 2026
Local Ai

FAST-GOAL: Fast and Efficient Global-local Object Alignment Learning

DGX agent

arXiv:2605.26615v1 Announce Type: new Abstract: Vision-language models such as CLIP have shown impressive capabilities in aligning images and text, but they often struggle with lengthy and detailed te

local-aiarxiv-cs-ai
27 May 2026
Model Releases

From PDF to RAG-Ready: Evaluating Document Conversion Frameworks for Domain-Specific Question Answering

DGX agent

arXiv:2604.04948v2 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) systems depend critically on the quality of document preprocessing, yet no prior study has evaluated PDF

model-releasesarxiv-cs-ai
27 May 2026
Model Releases

Helicase: Uncertainty-Guided Supply Chain Knowledge Graph Construction with Autonomous Multi-Agent LLMs

DGX agent

arXiv:2605.26835v1 Announce Type: new Abstract: LLM-based multi-agent systems have been widely adopted for knowledge retrieval and report generation, synthesizing known information through web search

model-releasesarxiv-cs-ai
27 May 2026
Safety

How Reliable are LLMs for Reasoning on the Re-ranking task?

DGX agent

arXiv:2508.18444v2 Announce Type: replace-cross Abstract: With the improving semantic understanding capability of Large Language Models (LLMs), they exhibit a greater awareness and alignment with huma

safetyarxiv-cs-ai
27 May 2026
Model Releases

HTMLCure: Turning Browser Experience into State Guided Repair for Interactive HTML

DGX agent

arXiv:2605.26807v1 Announce Type: cross Abstract: LLMs can now produce full HTML pages, but many of those pages are only superficially correct: they render once, then fail under scroll, hover, click,

model-releasesarxiv-cs-ai
27 May 2026
Applications

ICCU: In-Context Continual Unlearning via Pattern-Induced Refusal Rules

DGX agent

arXiv:2605.27138v1 Announce Type: new Abstract: Machine unlearning aims to remove the influence of specific data from trained language models. In real-world deployments, unlearning requests often arri

applicationsarxiv-cs-ai
27 May 2026
Model Releases

ImViD: Immersive Volumetric Videos for Enhanced VR Engagement

DGX agent

arXiv:2503.14359v2 Announce Type: replace Abstract: User engagement is greatly enhanced by fully immersive multi-modal experiences that combine visual and auditory stimuli. Consequently, the next fron

model-releasesarxiv-cs-cv
27 May 2026
Tutorials

In-Context Optimization for Retrieval-Augmented Generation: A Gradient-Descent Perspective

DGX agent

arXiv:2605.26356v1 Announce Type: new Abstract: In-context learning has recently been linked to implicit gradient descent in linear self-attention models, suggesting that context can induce a forward-

tutorialsarxiv-cs-cl
27 May 2026
Research

Innovation: An Almost Characterization of Hallucination

DGX agent

arXiv:2605.26808v1 Announce Type: cross Abstract: Hallucination is a central limitation of large language models (LLMs), and substantial effort has been devoted to understanding and mitigating it. Tow

researcharxiv-cs-ai
27 May 2026
Model Releases

Learnable Kernel Density Estimation for Graphs and Its Application to Graph-Level Anomaly Detection

DGX agent

arXiv:2505.21285v4 Announce Type: replace Abstract: This work proposes a framework LGKDE that learns kernel density estimation for graphs. The key challenge in graph density estimation lies in effecti

model-releasesarxiv-cs-lg
27 May 2026
Safety

Less is More: Early Stopping Rollout for On-Policy Distillation

DGX agent

arXiv:2605.27028v1 Announce Type: cross Abstract: On-policy distillation has recently emerged as a promising alternative to standard sequence-level imitation, training a student by scoring its own rol

safetyarxiv-cs-ai
27 May 2026
← Previous
1…682683684685686…1065
Next →