AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
50,764 results
Model Releases

PoisonForge: Task-Level Targeted Poisoning Benchmark for Instruction-Tuned LLMs

DGX agent

arXiv:2605.23168v1 Announce Type: cross Abstract: When practitioners fine-tune LLMs on unvetted datasets, an adversary can exploit the data supply chain through task-level poisoning: inserting a small

model-releasesarxiv-cs-ai
25 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Precise: SDE-Consistent Stochastic Sampling for RL Post-Training of Flow-Matching Models

DGX agent

arXiv:2605.23522v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become an effective way to improve prompt alignment and perceptual quality in diffusion and flow-matching generators.

safetyarxiv-cs-ai
25 May 2026
Research

SyMerge: From Non-Interference to Synergistic Merging via Single-Layer Adaptation

DGX agent

arXiv:2412.19098v4 Announce Type: replace Abstract: Model merging combines independently trained models into a single multi-task model. However, most existing approaches focus primarily on avoiding ta

researcharxiv-cs-lg
25 May 2026
Model Releases

Integrable Elasticity via Neural Demand Potentials

DGX agent

arXiv:2605.22820v1 Announce Type: new Abstract: We propose the Integrable Context-Dependent Demand Network (ICDN), a demand-first neural model for multiproduct retail demand. The model learns log-dema

model-releasesarxiv-cs-lg
23 May 2026
Model Releases

On Robustness and Chain-of-Thought Consistency of RL-Finetuned VLMs

DGX agent

arXiv:2602.12506v3 Announce Type: replace Abstract: Reinforcement learning (RL) finetuning has become a key technique for enhancing large language models (LLMs) on reasoning-intensive tasks, motivatin

model-releasesarxiv-cs-lg
23 May 2026
Research

Towards Solving the Gilbert-Pollak Conjecture via Large Language Models

DGX agent

arXiv:2601.22365v2 Announce Type: replace-cross Abstract: The Gilbert-Pollak Conjecture itep{gilbert1968steiner}, also known as the Steiner Ratio Conjecture, states that for any finite point set in th

researcharxiv-cs-lg
23 May 2026
Hardware

WarmServe: Enabling One-for-Many GPU Prewarming for Multi-LLM Serving

DGX agent

arXiv:2512.09472v2 Announce Type: replace-cross Abstract: Deploying multiple models within shared GPU clusters is a key strategy to improve resource efficiency in large language model (LLM) serving. E

hardwarearxiv-cs-lg
23 May 2026
Research

Access Paths for Efficient Ordering with Large Language Models

DGX agent

arXiv:2509.00303v3 Announce Type: replace-cross Abstract: In this work, we present the exttt{LLM ORDER BY} semantic operator as a logical abstraction and conduct a systematic study of its physical imp

researcharxiv-cs-ai
22 May 2026
Model Releases

Code Researcher: Deep Research Agent for Large Systems Code and Commit History

DGX agent

arXiv:2506.11060v2 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based coding agents have shown promising results on coding benchmarks, but their effectiveness on systems code rema

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

IdioLink: Retrieving Meaning Beyond Words Across Idiomatic and Literal Expressions

DGX agent

arXiv:2605.22247v1 Announce Type: new Abstract: Idioms pose a fundamental challenge for language models, as their meaning cannot be inferred from surface form alone. Understanding such expressions, th

model-releasesarxiv-cs-cl
22 May 2026
Safety

Modeling Emotional Dynamics in Agent-to-Agent Interactions on Moltbook

DGX agent

arXiv:2605.20442v1 Announce Type: cross Abstract: Generative AI systems are increasingly deployed as interactive agents in online environments, such as a social network called Moltbook. In Moltbook, l

safetyarxiv-cs-ai
22 May 2026
Model Releases

OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind

DGX agent

arXiv:2605.20423v1 Announce Type: new Abstract: Large Language Models (LLMs) perform well on many language tasks, but their Theory of Mind (ToM) reasoning is still uneven in complex social settings. E

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control

DGX agent

arXiv:2511.07820v3 Announce Type: replace-cross Abstract: Despite the rise of billion-parameter foundation models trained across thousands of GPUs, similar scaling gains have not been shown for humano

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Token-Level LLM Collaboration via FusionRoute

DGX agent

arXiv:2601.05106v4 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strengths across diverse domains. However, achieving strong performance across these domains with a singl

model-releasesarxiv-cs-cl
22 May 2026
Research

VISTA: Validation-Guided Integration of Spatial and Temporal Foundation Models with Anatomical Decoding for Rare-Pathology VCE Event Detection -- after competition results

DGX agent

arXiv:2605.22096v1 Announce Type: new Abstract: Capsule endoscopy event detection is challenging because clinically relevant findings are sparse, visually heterogeneous, and evaluated at the event lev

researcharxiv-cs-cv
22 May 2026
Model Releases

DriveMA: Rethinking Language Interfaces in Driving VLAs with One-Step Meta-Actions

DGX agent

arXiv:2605.21273v1 Announce Type: new Abstract: Driving Vision-Language-Action Models (Driving VLAs) commonly introduce natural-language reasoning as an intermediate interface for end-to-end planning,

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Findings of the Counter Turing Test: AI-Generated Text Detection

DGX agent

arXiv:2605.20761v1 Announce Type: new Abstract: The rapid proliferation of AI-generated text has introduced significant challenges in maintaining the integrity of digital content. Advanced generative

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining

DGX agent

arXiv:2605.20296v1 Announce Type: new Abstract: Fine-tuning a language model for a target task routinely degrades capabilities the training data never explicitly threatened. We study this phenomenon,

model-releasesarxiv-cs-lg
21 May 2026
Safety

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

DGX agent

arXiv:2506.08277v3 Announce Type: replace-cross Abstract: Recent voxel-wise multimodal brain encoding studies have shown that multimodal large language models (MLLMs) exhibit a higher degree of brain

safetyarxiv-cs-cl
21 May 2026
Model Releases

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs

DGX agent

arXiv:2505.19075v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable general capabilities, but enhancing skills such as reasoning often demands substanti

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

VISTAQA: Benchmarking Joint Visual Question Answering and Pixel-Level Evidence

DGX agent

arXiv:2605.20676v1 Announce Type: new Abstract: Establishing a clear link between model predictions and the visual evidence that supports them is critical for transparency and reliability in multimoda

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries

DGX agent

arXiv:2605.18891v1 Announce Type: cross Abstract: Evaluations of unlearning on reasoning models sometimes show a bypass pattern. The answer side looks unlearned, but the model's own thinking trace kee

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations

DGX agent

arXiv:2603.17305v2 Announce Type: replace Abstract: We propose CRAFT, a red-teaming alignment framework that leverages model reasoning capabilities and hidden representations to improve robustness aga

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Cross-View Attention Fusion Net: A Prior-Guided Dual-View Representation Learning for Cardiac Output Estimation from Short-Term PPG Signals

DGX agent

arXiv:2605.19666v1 Announce Type: cross Abstract: Accurate cardiac output (CO) estimation from photoplethysmography (PPG) is promising for unobtrusive hemodynamic monitoring, but remains difficult sin

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design

DGX agent

arXiv:2605.19743v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly applied to engineering design tasks, yet existing evaluation frameworks do not adequately address mul

model-releasesarxiv-cs-ai
20 May 2026
Agents

FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models

DGX agent

arXiv:2605.19111v1 Announce Type: cross Abstract: Existing text-to-image (T2I) evaluation metrics mainly assess whether generated images align with information explicitly stated in the prompt, but oft

agentsarxiv-cs-ai
20 May 2026
Local Ai

iDiff: Interpretable Difference-aware Framework for Pairwise Image Quality Assessment

DGX agent

arXiv:2605.19522v1 Announce Type: new Abstract: Pairwise image quality assessment (IQA) in professional photography requires a model not only to identify the preferred image between two candidates, bu

local-aiarxiv-cs-cv
20 May 2026
Safety

LambdaPO: A Lambda Style Policy Optimization for Reasoning Language Models

DGX agent

arXiv:2605.19416v1 Announce Type: new Abstract: Group Relative Policy Optimization(GRPO) has become a cornerstone of modern reinforcement learning alignment, prized for its efficacy in foregoing an ex

safetyarxiv-cs-cl
20 May 2026
Applications

LLM-MC-Affect: LLM-Based Monte Carlo Modeling of Affective Trajectories and Latent Ambiguity for Interpersonal Dynamic Insight

DGX agent

arXiv:2601.03645v2 Announce Type: replace Abstract: Emotional coordination is a core property of human interaction that shapes how relational meaning is constructed in real time. While text-based affe

applicationsarxiv-cs-cl
20 May 2026
Research

M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language Models

DGX agent

arXiv:2605.18774v1 Announce Type: cross Abstract: In long, multi-page industrial documents, retrieval-augmented generation (RAG) depends heavily on whether chunk boundaries follow the document's true

researcharxiv-cs-ai
20 May 2026
Model Releases

MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification

DGX agent

arXiv:2605.19359v1 Announce Type: new Abstract: Deep learning methods have demonstrated promising results in predicting BI-RADS scores from mammography images. However, the interpretation of these ima

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

POLAR-Bench: A Diagnostic Benchmark for Privacy-Utility Trade-offs in LLM Agents

DGX agent

arXiv:2605.19127v1 Announce Type: new Abstract: LLM agents increasingly have access to private user data and act on the user's behalf when interacting with third-party systems. The user defines what m

model-releasesarxiv-cs-ai
20 May 2026
Safety

Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models

DGX agent

arXiv:2605.19462v1 Announce Type: cross Abstract: The success of self-supervised learning (SSL) in vision and NLP has motivated its rapid adoption for time series. However, research has focused primar

safetyarxiv-cs-ai
20 May 2026
Research

Synergistic Foundation Models for Semi-Supervised Fetal Cardiac Ultrasound Analysis: SAM-Med2D Boundary Refinement and DINOv3 Semantic Enhancement

DGX agent

arXiv:2605.19799v1 Announce Type: cross Abstract: We present a semi-supervised framework for joint segmentation and classification of fetal cardiac ultrasound images. Built upon the EchoCare multi-tas

researcharxiv-cs-ai
20 May 2026
Model Releases

TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards

DGX agent

arXiv:2605.19320v1 Announce Type: new Abstract: Faithful text rendering remains a persistent weakness of large text-to-image generative models, as it requires both semantic instruction following and f

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

The Evaluation Game: Beyond Static LLM Benchmarking

DGX agent

arXiv:2605.19377v1 Announce Type: cross Abstract: As jailbreaks, adversarially crafted inputs that bypass safety constraints, continue to be discovered in Large Language Models, practitioners increasi

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

To Call or Not to Call: Diagnosing Intrinsic Over-Calling Bias in LLM Agents

DGX agent

arXiv:2605.18882v1 Announce Type: cross Abstract: LLM agents exhibit a consistent tendency to over-call, invoking tools even in situations where none is needed. On the When2Call benchmark, six models

model-releasesarxiv-cs-ai
20 May 2026
Hardware

UCCI: Calibrated Uncertainty for Cost-Optimal LLM Cascade Routing

DGX agent

arXiv:2605.18796v1 Announce Type: cross Abstract: LLM cascades and model routing promise lower inference cost by sending easy queries to a small model and escalating hard ones to a large model, but mo

hardwarearxiv-cs-cl
20 May 2026
Agents

VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving

DGX agent

arXiv:2605.20082v1 Announce Type: cross Abstract: The rapid growth of autonomous driving datasets has enabled the scaling of powerful motion forecasting models. While large-scale pretraining provides

agentsarxiv-cs-ai
20 May 2026
Safety

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making

DGX agent

arXiv:2602.07008v2 Announce Type: replace Abstract: Reliable models should not only predict correctly, but also justify decisions with acceptable evidence. Yet conventional supervised learning typical

safetyarxiv-cs-cv
20 May 2026
Model Releases

Conv-FinRe: A Conversational and Longitudinal Benchmark for Utility-Grounded Financial Recommendation

DGX agent

arXiv:2602.16990v2 Announce Type: replace Abstract: Most recommendation benchmarks evaluate how well a model imitates user behavior. In financial advisory, however, observed actions can be noisy or sh

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Detecting Verbatim LLM Copy-Paste in Homework

DGX agent

arXiv:2605.16336v1 Announce Type: cross Abstract: Large language models (LLMs) have made fluent essay writing, code drafting, and quiz answering instantly available to students at every level, from se

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Disentangled Latent Dynamics Manifold Fusion for Solving Parameterized PDEs

DGX agent

arXiv:2603.12676v2 Announce Type: replace Abstract: Generalizing neural surrogate models across different PDE parameters remains difficult because changes in PDE coefficients often make learning harde

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control

DGX agent

arXiv:2605.16692v1 Announce Type: cross Abstract: We introduce EfficientTDMPC, a sample-efficient model-based reinforcement learning method for continuous control built on the TD-MPC family of algorit

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Episodic-Semantic Memory Architecture for Long-Horizon Scientific Agents

DGX agent

arXiv:2605.17625v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into persistent scientific collaborators, context window saturation has emerged as a critical bottleneck. Scienti

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GLT-PEFT: Gated Lie-Tucker Parameter-Efficient Fine-Tuning for Alzheimer's Disease Diagnosis with Hippocampal Segmentation Pretraining

DGX agent

arXiv:2605.16769v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) has emerged as a promising paradigm for adapting pretrained models under limited data conditions. However, most e

model-releasesarxiv-cs-cv
19 May 2026
Research

Language Models as Efficient Reward Function Searchers for Custom-Environment Multi-Objective Reinforcement

DGX agent

arXiv:2409.02428v4 Announce Type: replace-cross Abstract: Achieving the effective design and improvement of reward functions in reinforcement learning (RL) tasks with complex custom environments and m

researcharxiv-cs-ai
19 May 2026
Hardware

LEAP: Learnable End-to-End Adaptive Pruning of Large Language Models

DGX agent

arXiv:2605.17289v1 Announce Type: cross Abstract: Unstructured sparsity is now natively accelerated by recent GPU kernels and dataflow hardware, shifting the bottleneck from inference execution to the

hardwarearxiv-cs-ai
19 May 2026
← Previous
1…288289290291292…1058
Next →