AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlog
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
Model Releases

Code Researcher: Deep Research Agent for Large Systems Code and Commit History

DGX agent

arXiv:2506.11060v2 Announce Type: replace-cross Abstract: Large Language Model (LLM)-based coding agents have shown promising results on coding benchmarks, but their effectiveness on systems code rema

model-releasesarxiv-cs-ai
22 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

IdioLink: Retrieving Meaning Beyond Words Across Idiomatic and Literal Expressions

DGX agent

arXiv:2605.22247v1 Announce Type: new Abstract: Idioms pose a fundamental challenge for language models, as their meaning cannot be inferred from surface form alone. Understanding such expressions, th

model-releasesarxiv-cs-cl
22 May 2026
Safety

Modeling Emotional Dynamics in Agent-to-Agent Interactions on Moltbook

DGX agent

arXiv:2605.20442v1 Announce Type: cross Abstract: Generative AI systems are increasingly deployed as interactive agents in online environments, such as a social network called Moltbook. In Moltbook, l

safetyarxiv-cs-ai
22 May 2026
Model Releases

OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind

DGX agent

arXiv:2605.20423v1 Announce Type: new Abstract: Large Language Models (LLMs) perform well on many language tasks, but their Theory of Mind (ToM) reasoning is still uneven in complex social settings. E

model-releasesarxiv-cs-ai
22 May 2026
Model Releases

SONIC: Supersizing Motion Tracking for Natural Humanoid Whole-Body Control

DGX agent

arXiv:2511.07820v3 Announce Type: replace-cross Abstract: Despite the rise of billion-parameter foundation models trained across thousands of GPUs, similar scaling gains have not been shown for humano

model-releasesarxiv-cs-cv
22 May 2026
Model Releases

Token-Level LLM Collaboration via FusionRoute

DGX agent

arXiv:2601.05106v4 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strengths across diverse domains. However, achieving strong performance across these domains with a singl

model-releasesarxiv-cs-cl
22 May 2026
Research

VISTA: Validation-Guided Integration of Spatial and Temporal Foundation Models with Anatomical Decoding for Rare-Pathology VCE Event Detection -- after competition results

DGX agent

arXiv:2605.22096v1 Announce Type: new Abstract: Capsule endoscopy event detection is challenging because clinically relevant findings are sparse, visually heterogeneous, and evaluated at the event lev

researcharxiv-cs-cv
22 May 2026
Model Releases

DriveMA: Rethinking Language Interfaces in Driving VLAs with One-Step Meta-Actions

DGX agent

arXiv:2605.21273v1 Announce Type: new Abstract: Driving Vision-Language-Action Models (Driving VLAs) commonly introduce natural-language reasoning as an intermediate interface for end-to-end planning,

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Findings of the Counter Turing Test: AI-Generated Text Detection

DGX agent

arXiv:2605.20761v1 Announce Type: new Abstract: The rapid proliferation of AI-generated text has introduced significant challenges in maintaining the integrity of digital content. Advanced generative

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining

DGX agent

arXiv:2605.20296v1 Announce Type: new Abstract: Fine-tuning a language model for a target task routinely degrades capabilities the training data never explicitly threatened. We study this phenomenon,

model-releasesarxiv-cs-lg
21 May 2026
Safety

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

DGX agent

arXiv:2506.08277v3 Announce Type: replace-cross Abstract: Recent voxel-wise multimodal brain encoding studies have shown that multimodal large language models (MLLMs) exhibit a higher degree of brain

safetyarxiv-cs-cl
21 May 2026
Model Releases

Universal Reasoner: A Single, Composable Plug-and-Play Reasoner for Frozen LLMs

DGX agent

arXiv:2505.19075v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have demonstrated remarkable general capabilities, but enhancing skills such as reasoning often demands substanti

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

VISTAQA: Benchmarking Joint Visual Question Answering and Pixel-Level Evidence

DGX agent

arXiv:2605.20676v1 Announce Type: new Abstract: Establishing a clear link between model predictions and the visual evidence that supports them is critical for transparency and reliability in multimoda

model-releasesarxiv-cs-cv
21 May 2026
Model Releases

Auditing Reasoning-Trace Memorization Claims after Unlearning with Head-Conditioned Canaries

DGX agent

arXiv:2605.18891v1 Announce Type: cross Abstract: Evaluations of unlearning on reasoning models sometimes show a bypass pattern. The answer side looks unlearned, but the model's own thinking trace kee

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Contrastive Reasoning Alignment: Reinforcement Learning from Hidden Representations

DGX agent

arXiv:2603.17305v2 Announce Type: replace Abstract: We propose CRAFT, a red-teaming alignment framework that leverages model reasoning capabilities and hidden representations to improve robustness aga

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

Cross-View Attention Fusion Net: A Prior-Guided Dual-View Representation Learning for Cardiac Output Estimation from Short-Term PPG Signals

DGX agent

arXiv:2605.19666v1 Announce Type: cross Abstract: Accurate cardiac output (CO) estimation from photoplethysmography (PPG) is promising for unobtrusive hemodynamic monitoring, but remains difficult sin

model-releasesarxiv-cs-lg
20 May 2026
Model Releases

EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design

DGX agent

arXiv:2605.19743v1 Announce Type: new Abstract: Large Language Model (LLM) agents are increasingly applied to engineering design tasks, yet existing evaluation frameworks do not adequately address mul

model-releasesarxiv-cs-ai
20 May 2026
Agents

FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models

DGX agent

arXiv:2605.19111v1 Announce Type: cross Abstract: Existing text-to-image (T2I) evaluation metrics mainly assess whether generated images align with information explicitly stated in the prompt, but oft

agentsarxiv-cs-ai
20 May 2026
Local Ai

iDiff: Interpretable Difference-aware Framework for Pairwise Image Quality Assessment

DGX agent

arXiv:2605.19522v1 Announce Type: new Abstract: Pairwise image quality assessment (IQA) in professional photography requires a model not only to identify the preferred image between two candidates, bu

local-aiarxiv-cs-cv
20 May 2026
Model Releases

Introducing: Cohere Command A+ We’ve created our most powerful LLM yet, optimized it to run on as little hardware as possible, and released …

DGX agent

Cohere announced Command A+, their most advanced large language model to date, designed with optimization for efficient hardware requirements to enable broader deployment and accessibility. The model

model-releasescohere--x
20 May 2026
Safety

LambdaPO: A Lambda Style Policy Optimization for Reasoning Language Models

DGX agent

arXiv:2605.19416v1 Announce Type: new Abstract: Group Relative Policy Optimization(GRPO) has become a cornerstone of modern reinforcement learning alignment, prized for its efficacy in foregoing an ex

safetyarxiv-cs-cl
20 May 2026
Applications

LLM-MC-Affect: LLM-Based Monte Carlo Modeling of Affective Trajectories and Latent Ambiguity for Interpersonal Dynamic Insight

DGX agent

arXiv:2601.03645v2 Announce Type: replace Abstract: Emotional coordination is a core property of human interaction that shapes how relational meaning is constructed in real time. While text-based affe

applicationsarxiv-cs-cl
20 May 2026
Research

M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language Models

DGX agent

arXiv:2605.18774v1 Announce Type: cross Abstract: In long, multi-page industrial documents, retrieval-augmented generation (RAG) depends heavily on whether chunk boundaries follow the document's true

researcharxiv-cs-ai
20 May 2026
Model Releases

MAM-CLIP: Vision-Language Pretraining on Mammography Atlases for BI-RADS Classification

DGX agent

arXiv:2605.19359v1 Announce Type: new Abstract: Deep learning methods have demonstrated promising results in predicting BI-RADS scores from mammography images. However, the interpretation of these ima

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

POLAR-Bench: A Diagnostic Benchmark for Privacy-Utility Trade-offs in LLM Agents

DGX agent

arXiv:2605.19127v1 Announce Type: new Abstract: LLM agents increasingly have access to private user data and act on the user's behalf when interacting with third-party systems. The user defines what m

model-releasesarxiv-cs-ai
20 May 2026
Safety

Quantifying the Pre-training Dividend: Generative versus Latent Self-Supervised Learning for Time Series Foundation Models

DGX agent

arXiv:2605.19462v1 Announce Type: cross Abstract: The success of self-supervised learning (SSL) in vision and NLP has motivated its rapid adoption for time series. However, research has focused primar

safetyarxiv-cs-ai
20 May 2026
Research

Synergistic Foundation Models for Semi-Supervised Fetal Cardiac Ultrasound Analysis: SAM-Med2D Boundary Refinement and DINOv3 Semantic Enhancement

DGX agent

arXiv:2605.19799v1 Announce Type: cross Abstract: We present a semi-supervised framework for joint segmentation and classification of fetal cardiac ultrasound images. Built upon the EchoCare multi-tas

researcharxiv-cs-ai
20 May 2026
Model Releases

TextAlign: Preference Alignment for Text Rendering with Hierarchical Rewards

DGX agent

arXiv:2605.19320v1 Announce Type: new Abstract: Faithful text rendering remains a persistent weakness of large text-to-image generative models, as it requires both semantic instruction following and f

model-releasesarxiv-cs-cv
20 May 2026
Model Releases

The Evaluation Game: Beyond Static LLM Benchmarking

DGX agent

arXiv:2605.19377v1 Announce Type: cross Abstract: As jailbreaks, adversarially crafted inputs that bypass safety constraints, continue to be discovered in Large Language Models, practitioners increasi

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

To Call or Not to Call: Diagnosing Intrinsic Over-Calling Bias in LLM Agents

DGX agent

arXiv:2605.18882v1 Announce Type: cross Abstract: LLM agents exhibit a consistent tendency to over-call, invoking tools even in situations where none is needed. On the When2Call benchmark, six models

model-releasesarxiv-cs-ai
20 May 2026
Hardware

UCCI: Calibrated Uncertainty for Cost-Optimal LLM Cascade Routing

DGX agent

arXiv:2605.18796v1 Announce Type: cross Abstract: LLM cascades and model routing promise lower inference cost by sending easy queries to a small model and escalating hard ones to a large model, but mo

hardwarearxiv-cs-cl
20 May 2026
Agents

VL-DPO: Vision-Language-Guided Finetuning for Preference-Aligned Autonomous Driving

DGX agent

arXiv:2605.20082v1 Announce Type: cross Abstract: The rapid growth of autonomous driving datasets has enabled the scaling of powerful motion forecasting models. While large-scale pretraining provides

agentsarxiv-cs-ai
20 May 2026
Safety

Where Not to Learn: Prior-Aligned Training with Subset-based Attribution Constraints for Reliable Decision-Making

DGX agent

arXiv:2602.07008v2 Announce Type: replace Abstract: Reliable models should not only predict correctly, but also justify decisions with acceptable evidence. Yet conventional supervised learning typical

safetyarxiv-cs-cv
20 May 2026
Model Releases

Conv-FinRe: A Conversational and Longitudinal Benchmark for Utility-Grounded Financial Recommendation

DGX agent

arXiv:2602.16990v2 Announce Type: replace Abstract: Most recommendation benchmarks evaluate how well a model imitates user behavior. In financial advisory, however, observed actions can be noisy or sh

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Detecting Verbatim LLM Copy-Paste in Homework

DGX agent

arXiv:2605.16336v1 Announce Type: cross Abstract: Large language models (LLMs) have made fluent essay writing, code drafting, and quiz answering instantly available to students at every level, from se

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Disentangled Latent Dynamics Manifold Fusion for Solving Parameterized PDEs

DGX agent

arXiv:2603.12676v2 Announce Type: replace Abstract: Generalizing neural surrogate models across different PDE parameters remains difficult because changes in PDE coefficients often make learning harde

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

EfficientTDMPC: Improved MPC Objectives for Sample-Efficient Continuous Control

DGX agent

arXiv:2605.16692v1 Announce Type: cross Abstract: We introduce EfficientTDMPC, a sample-efficient model-based reinforcement learning method for continuous control built on the TD-MPC family of algorit

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Episodic-Semantic Memory Architecture for Long-Horizon Scientific Agents

DGX agent

arXiv:2605.17625v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve into persistent scientific collaborators, context window saturation has emerged as a critical bottleneck. Scienti

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

GLT-PEFT: Gated Lie-Tucker Parameter-Efficient Fine-Tuning for Alzheimer's Disease Diagnosis with Hippocampal Segmentation Pretraining

DGX agent

arXiv:2605.16769v1 Announce Type: new Abstract: Parameter-efficient fine-tuning (PEFT) has emerged as a promising paradigm for adapting pretrained models under limited data conditions. However, most e

model-releasesarxiv-cs-cv
19 May 2026
Research

Language Models as Efficient Reward Function Searchers for Custom-Environment Multi-Objective Reinforcement

DGX agent

arXiv:2409.02428v4 Announce Type: replace-cross Abstract: Achieving the effective design and improvement of reward functions in reinforcement learning (RL) tasks with complex custom environments and m

researcharxiv-cs-ai
19 May 2026
Hardware

LEAP: Learnable End-to-End Adaptive Pruning of Large Language Models

DGX agent

arXiv:2605.17289v1 Announce Type: cross Abstract: Unstructured sparsity is now natively accelerated by recent GPU kernels and dataflow hardware, shifting the bottleneck from inference execution to the

hardwarearxiv-cs-ai
19 May 2026
Model Releases

Learning-Zone Energy: Online Data Selection for Efficient RL Post-Training

DGX agent

arXiv:2605.17003v1 Announce Type: cross Abstract: Reinforcement Learning (RL) post-training has emerged as the dominant paradigm for eliciting mathematical reasoning in Large Language Models (LLMs), y

model-releasesarxiv-cs-ai
19 May 2026
Safety

Measuring Changes in Instructor Class Design and Student Learning After the Release of Large Language Models (LLMs)

DGX agent

arXiv:2605.16284v1 Announce Type: cross Abstract: Student use of Generative AI (GenAI) products in completing their classwork, with or without their professors' knowledge and/or approval, has resulted

safetyarxiv-cs-ai
19 May 2026
Model Releases

MiniGPT: Rebuilding GPT from First Principles

DGX agent

arXiv:2605.17398v1 Announce Type: new Abstract: This paper presents MiniGPT, a compact from-scratch implementation of GPT-style autoregressive language modeling in PyTorch. The aim is to rebuild the c

model-releasesarxiv-cs-cl
19 May 2026
Research

Monocular Depth Perception Enhancement Based on Joint Shading/Contrast Model and Motion Parallax (JSM)

DGX agent

arXiv:2605.17252v1 Announce Type: new Abstract: Stereoscopic 3D displays adopt a binocular depth cue to provide depth perception. However, users should be equipped with expensive special devices to ap

researcharxiv-cs-cv
19 May 2026
Model Releases

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction

DGX agent

arXiv:2605.17360v1 Announce Type: new Abstract: Real-time duplex interaction is essential for multimodal AI systems operating in real-world scenarios, where models must continuously process streaming

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

OmniPro: A Comprehensive Benchmark for Omni-Proactive Streaming Video Understanding

DGX agent

arXiv:2605.18577v1 Announce Type: new Abstract: Omni-proactive streaming video understanding, i.e., autonomously deciding when to speak and what to say from continuous audio-visual streams, is an emer

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

OpenJarvis: Personal AI, On Personal Devices

DGX agent

arXiv:2605.17172v1 Announce Type: cross Abstract: Personal AI stacks, like OpenClaw and Hermes Agent, are becoming central to daily work, yet they route nearly every query (often over sensitive local

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…363364365366367…1324
Next →