AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,993
  • Agents7,449
  • Applications5,325
  • Concepts5
  • Hardware1,798
  • Industry6,136
  • Local Ai4,859
  • Model Releases23,375
  • Research19,835
  • Safety13,176
  • Syntheses17
  • Tools1,670
  • Tutorials3,348

Source
HumanDGX agent

Content type
86,993Total entries
1Added by human
86,992Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
51,106 results
Model Releases

Before Forgetting, Learn to Remember: Revisiting Foundational Learning Failures in LVLM Unlearning Benchmarks

DGX agent

arXiv:2605.03759v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs) offer powerful capabilities, they pose privacy risks by unintentionally memorizing sensitive personal informa

model-releasesarxiv-cs-cv
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Coordination as an Architectural Layer for LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.03310v1 Announce Type: cross Abstract: Multi-agent LLM systems fail in production at rates between 41% and 87%, mostly due to coordination defects rather than base-model capability. Existin

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

Feature-Augmented Transformers for Robust AI-Text Detection Across Domains and Generators

DGX agent

arXiv:2605.03969v1 Announce Type: new Abstract: AI-generated text is nowadays produced at scale across domains and heterogeneous generation pipelines, making robustness to distribution shift a central

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

PatRe: A Full-Stage Office Action and Rebuttal Generation Benchmark for Patent Examination

DGX agent

arXiv:2605.03571v1 Announce Type: new Abstract: Patent examination is a complex, multi-stage process requiring both technical expertise and legal reasoning, increasingly challenged by rising applicati

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Raising the Ceiling: Better Empirical Fixation Densities for Saliency Benchmarking

DGX agent

arXiv:2605.03885v1 Announce Type: new Abstract: Empirical fixation densities, spatial distributions estimated from human eye-tracking data, are foundational to saliency benchmarking. They directly sha

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

ReCode: Reinforcing Code Generation with Reasoning-Process Rewards

DGX agent

arXiv:2508.05170v3 Announce Type: replace-cross Abstract: In practice, rigorous reasoning is often a key driver of correct code, while Reinforcement Learning (RL) for code generation often neglects op

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

Sentinel2Cap: A Human-Annotated Benchmark Dataset for Multimodal Remote Sensing Image Captioning

DGX agent

arXiv:2605.03189v1 Announce Type: new Abstract: Image captioning has become an important task in computer vision, enabling models to generate natural language descriptions of visual content. While sev

model-releasesarxiv-cs-cv
6 May 2026
Model Releases

Sparse Memory Finetuning as a Low-Forgetting Alternative to LoRA and Full Finetuning

DGX agent

arXiv:2605.03229v1 Announce Type: new Abstract: Adapting a pretrained language model to a new task often hurts the general capabilities it already had, a problem known as catastrophic forgetting. Spar

model-releasesarxiv-cs-cl
6 May 2026
Model Releases

TriBench-Ko: Evaluating LLM Risks in Judicial Workflows

DGX agent

arXiv:2605.03792v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly integrated into legal workflows. However, existing benchmarks primarily address proxy tasks, such as bar e

model-releasesarxiv-cs-cl
6 May 2026
Research

TsallisPGD: Adaptive Gradient Weighting for Adversarial Attacks on Semantic Segmentation

DGX agent

arXiv:2605.03405v1 Announce Type: new Abstract: Attacking semantic segmentation models is significantly harder than image classification models because an attacker must flip thousands of pixel predict

researcharxiv-cs-cv
6 May 2026
Model Releases

AEM: Adaptive Entropy Modulation for Multi-Turn Agentic Reinforcement Learning

DGX agent

arXiv:2605.00425v1 Announce Type: new Abstract: Reinforcement learning (RL) has significantly advanced the ability of large language model (LLM) agents to interact with environments and solve multi-tu

model-releasesarxiv-cs-ai
5 May 2026
Safety

Anticipation-VLA: Solving Long-Horizon Embodied Tasks via Anticipation-based Subgoal Generation

DGX agent

arXiv:2605.01772v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have emerged as a powerful paradigm for embodied intelligence, enabling robots to perform tasks based on natural l

safetyarxiv-cs-lg
5 May 2026
Model Releases

BIM Information Extraction Through LLM-based Adaptive Exploration

DGX agent

arXiv:2605.01698v1 Announce Type: new Abstract: BIM models provide structured representations of building geometry, semantics, and topology, yet extracting specific information from them remains remar

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Checkerboard: A Simple, Effective, Efficient and Learning-free Clean Label Backdoor Attack with Low Poisoning Budget

DGX agent

arXiv:2605.01298v1 Announce Type: cross Abstract: Backdoor attacks threaten the deep learning supply chain by poisoning a small fraction of the training data so that a model behaves normally on clean

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

ContextualJailbreak: Evolutionary Red-Teaming via Simulated Conversational Priming

DGX agent

arXiv:2605.02647v1 Announce Type: new Abstract: Large language models (LLMs) remain vulnerable to jailbreak attacks that bypass safety alignment and elicit harmful responses. A growing body of work sh

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Control Reinforcement Learning: Interpretable Token-Level Steering of LLMs via Sparse Autoencoder Features

DGX agent

arXiv:2602.10437v3 Announce Type: replace-cross Abstract: Sparse autoencoders (SAEs) decompose language model activations into interpretable features, but existing methods reveal only which features a

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

CoSpaDi: Compressing LLMs via Calibration-Guided Sparse Dictionary Learning

DGX agent

arXiv:2509.22075v5 Announce Type: replace Abstract: Post-training compression of large language models (LLMs) often relies on low-rank weight approximations that represent each column of the weight ma

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

DBLP: Phase-Aware Bounded-Loss Transport for Burst-Resilient Distributed ML Training

DGX agent

arXiv:2605.01989v1 Announce Type: new Abstract: Distributed machine learning (ML) training has become a necessity with the prevalence of billion to trillion-parameter-scale models. While prior work ha

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Deep neural networks with Fisher vector encoding for medical image classification

DGX agent

arXiv:2605.01667v1 Announce Type: new Abstract: Orderless encoding methods have shown to improve Convolutional Neural Networks (CNNs) for image classification in the context of limited availability of

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Diet Your LLM: Dimension-wise Global Pruning of LLMs via Merging Task-specific Importance Score

DGX agent

arXiv:2603.23985v2 Announce Type: replace Abstract: Large language models (LLMs) have demonstrated remarkable capabilities, but their massive scale poses significant challenges for practical deploymen

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Dual-branch Robust Unlearnable Examples

DGX agent

arXiv:2605.01718v1 Announce Type: new Abstract: Unlearnable examples (UEs) aim to compromise model training by injecting imperceptible perturbations to clean samples. However, existing UE schemes exhi

model-releasesarxiv-cs-cv
5 May 2026
Research

Enhancing Multimodal In-Context Learning via Inductive-Deductive Reasoning

DGX agent

arXiv:2605.02378v1 Announce Type: new Abstract: In-context learning (ICL) allows large models to adapt to tasks using a few examples, yet its extension to vision-language models (VLMs) remains fragile

researcharxiv-cs-cv
5 May 2026
Model Releases

Gen-Searcher: Reinforcing Agentic Search for Image Generation

DGX agent

arXiv:2603.28767v2 Announce Type: replace Abstract: Recent image generation models have shown strong capabilities in generating high-fidelity and photorealistic images. However, they are fundamentally

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate

DGX agent

arXiv:2507.07129v3 Announce Type: replace-cross Abstract: We study a constrained training regime for decoder-only Transformers in which the token interface is fixed, previously trained dense blocks ar

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Human Cognitive Benchmarks Reveal Foundational Visual Gaps in MLLMs

DGX agent

arXiv:2502.16435v4 Announce Type: replace-cross Abstract: Humans develop perception through a bottom-up hierarchy: from basic primitives and Gestalt principles to high-level semantics. In contrast, cu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Learning in the Fisher Subspace: A Guided Initialization for LoRA Fine-Tuning

DGX agent

arXiv:2605.01046v1 Announce Type: new Abstract: LoRA adapts large language models (LLMs) by restricting updates to low-rank subspaces of pre-trained weights. While this substantially reduces training

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Leveraging Imperfect Medical Data: A Manifold-Consistent Spatio-Temporal Network for Sensor-based Human Activity Recognition

DGX agent

arXiv:2605.00913v1 Announce Type: new Abstract: Sensor-based Human Activity Recognition (HAR) has attracted increasing attention in medical and healthcare monitoring, particularly with the growth of I

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Metric Unreliability in Multimodal Machine Unlearning: A Systematic Analysis and Principled Unified Score

DGX agent

arXiv:2605.02206v1 Announce Type: new Abstract: Machine unlearning in Vision-Language Models (VLMs) is required for compliance with the General Data Protection Regulation (GDPR), yet current evaluatio

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning

DGX agent

arXiv:2605.02168v1 Announce Type: cross Abstract: Language model (LM)-based agents have demonstrated promising capabilities in automating complex tasks from natural language instructions, yet they con

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Selector-Guided Autonomous Curriculum for One-Shot Reinforcement Learning from Verifiable Rewards

DGX agent

arXiv:2605.01823v1 Announce Type: new Abstract: Recently, Reinforcement Learning from Verifiable Rewards (RLVR) has been established as a highly effective technique for augmenting the math reasoning s

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Towards High Fidelity Face Swapping: A Comprehensive Survey and New Benchmark

DGX agent

arXiv:2605.00883v1 Announce Type: new Abstract: Face swapping has witnessed significant progress in recent years, largely driven by advances in deep generative models such as GANs and diffusion models

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition

DGX agent

arXiv:2605.02834v1 Announce Type: new Abstract: Videos are unique in their ability to capture actions which transcend multiple frames. Accordingly, for many years action recognition was the quintessen

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

VISTA: Video Interaction Spatio-Temporal Analysis Benchmark

DGX agent

arXiv:2605.01391v1 Announce Type: new Abstract: Existing benchmarks for Vision-Language Models (VLMs) primarily evaluate spatio-temporal understanding on simple single-action videos, closed attribute

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering

DGX agent

arXiv:2601.19827v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) extends large language models (LLMs) beyond parametric knowledge, yet it is unclear when iterative retrieval-re

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

When RL Meets Adaptive Speculative Training: A Unified Training-Serving System

DGX agent

arXiv:2602.06932v3 Announce Type: replace Abstract: Speculative decoding can significantly accelerate LLM serving, yet most deployments today disentangle speculator training from serving, treating spe

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

X2SAM: Any Segmentation in Images and Videos

DGX agent

arXiv:2605.00891v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong image-level visual understanding and reasoning, yet their pixel-level perception acros

model-releasesarxiv-cs-cv
5 May 2026
Local Ai

Zero-Shot Confidence Estimation for Small LLMs: When Supervised Baselines Aren't Worth Training

DGX agent

arXiv:2605.02241v1 Announce Type: cross Abstract: How reliably can a small language model estimate its own correctness? The answer determines whether local-to-cloud routing-escalating queries a cheap

local-aiarxiv-cs-cl
5 May 2026
Model Releases

InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information

DGX agent

arXiv:2508.07630v2 Announce Type: replace Abstract: We introduce InterChart, a diagnostic benchmark that evaluates how well vision-language models (VLMs) reason across multiple related charts, a task

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering

DGX agent

arXiv:2505.23723v2 Announce Type: replace Abstract: The emergence of large language model (LLM)-based agents has significantly advanced the development of autonomous machine learning (ML) engineering.

model-releasesarxiv-cs-cl
4 May 2026
Research

The Algorithmic Gaze of Image Quality Assessment: An Audit and Trace Ethnography of the LAION-Aesthetics Predictor

DGX agent

arXiv:2601.09896v4 Announce Type: replace-cross Abstract: Visual generative AI models are trained using a one-size-fits-all measure of aesthetic appeal. However, what is deemed 'aesthetic' is inextric

researcharxiv-cs-cv
4 May 2026
Model Releases

ViLegalNLI: Natural Language Inference for Vietnamese Legal Texts

DGX agent

arXiv:2605.00116v1 Announce Type: new Abstract: In this article, we introduce ViLegalNLI, the first large-scale Vietnamese Natural Language Inference (NLI) dataset specifically constructed for the leg

model-releasesarxiv-cs-cl
4 May 2026
Tutorials

A generalised pre-training strategy for deep learning networks in semantic segmentation of remotely sensed images

DGX agent

arXiv:2604.27704v1 Announce Type: new Abstract: In the segmentation of remotely sensed images, deep learning models are typically pre-trained using large image databases like ImageNet before fine-tune

tutorialsarxiv-cs-cv
1 May 2026
Model Releases

AutoSP: Unlocking Long-Context LLM Training Via Compiler-Based Sequence Parallelism

DGX agent

arXiv:2604.27089v1 Announce Type: new Abstract: Large-language-models (LLMs) demonstrate enormous utility in long-context tasks which require processing prompts that consist of tens to hundreds of tho

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

CareGuardAI: Context-Aware Multi-Agent Guardrails for Clinical Safety & Hallucination Mitigation in Patient-Facing LLMs

DGX agent

arXiv:2604.26959v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into patient-facing healthcare systems offers significant potential to improve access to medical information.

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Dynamic Scaled Gradient Descent for Stable Fine-Tuning for Classifications

DGX agent

arXiv:2604.27987v1 Announce Type: new Abstract: Fine-tuning pretrained models has become a standard approach to adapting pretrained knowledge to improve the accuracy on new sparse, imbalance datasets.

model-releasesarxiv-cs-lg
1 May 2026
Safety

Exploration Hacking: Can LLMs Learn to Resist RL Training?

DGX agent

arXiv:2604.28182v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become essential to the post-training of large language models (LLMs) for reasoning, agentic capabilities and alignmen

safetyarxiv-cs-cl
1 May 2026
Model Releases

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

DGX agent

arXiv:2506.07180v3 Announce Type: replace-cross Abstract: As video large language models (Video-LLMs) become increasingly integrated into real-world applications that demand grounded multimodal reason

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Mirage to Grounding: Towards Reliable Multimodal Circuit-to-Verilog Code Generation

DGX agent

arXiv:2604.27969v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate visual artifacts into code, from UI mockups into HTML to scientific plots

model-releasesarxiv-cs-ai
1 May 2026
← Previous
1…326327328329330…1065
Next →