AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,766 results
Model Releases

Dual-branch Robust Unlearnable Examples

DGX agent

arXiv:2605.01718v1 Announce Type: new Abstract: Unlearnable examples (UEs) aim to compromise model training by injecting imperceptible perturbations to clean samples. However, existing UE schemes exhi

model-releasesarxiv-cs-cv
5 May 2026
Research

Enhancing Multimodal In-Context Learning via Inductive-Deductive Reasoning

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.02378v1 Announce Type: new Abstract: In-context learning (ICL) allows large models to adapt to tasks using a few examples, yet its extension to vision-language models (VLMs) remains fragile

researcharxiv-cs-cv
5 May 2026
Model Releases

Gen-Searcher: Reinforcing Agentic Search for Image Generation

DGX agent

arXiv:2603.28767v2 Announce Type: replace Abstract: Recent image generation models have shown strong capabilities in generating high-fidelity and photorealistic images. However, they are fundamentally

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

GPT-5.5 Instant: smarter, clearer, and more personalized

DGX agent

GPT-5.5 Instant is OpenAI's faster, more efficient variant of their GPT-5.5 model, designed to deliver improved reasoning and clarity while maintaining lower latency for real-time applications. The mo

model-releasesopenai
5 May 2026
Model Releases

Growing Transformers: Modular Composition and Layer-wise Expansion on a Frozen Substrate

DGX agent

arXiv:2507.07129v3 Announce Type: replace-cross Abstract: We study a constrained training regime for decoder-only Transformers in which the token interface is fixed, previously trained dense blocks ar

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Human Cognitive Benchmarks Reveal Foundational Visual Gaps in MLLMs

DGX agent

arXiv:2502.16435v4 Announce Type: replace-cross Abstract: Humans develop perception through a bottom-up hierarchy: from basic primitives and Gestalt principles to high-level semantics. In contrast, cu

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Learning in the Fisher Subspace: A Guided Initialization for LoRA Fine-Tuning

DGX agent

arXiv:2605.01046v1 Announce Type: new Abstract: LoRA adapts large language models (LLMs) by restricting updates to low-rank subspaces of pre-trained weights. While this substantially reduces training

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Leveraging Imperfect Medical Data: A Manifold-Consistent Spatio-Temporal Network for Sensor-based Human Activity Recognition

DGX agent

arXiv:2605.00913v1 Announce Type: new Abstract: Sensor-based Human Activity Recognition (HAR) has attracted increasing attention in medical and healthcare monitoring, particularly with the growth of I

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Metric Unreliability in Multimodal Machine Unlearning: A Systematic Analysis and Principled Unified Score

DGX agent

arXiv:2605.02206v1 Announce Type: new Abstract: Machine unlearning in Vision-Language Models (VLMs) is required for compliance with the General Data Protection Regulation (GDPR), yet current evaluatio

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

Planner Matters! An Efficient and Unbalanced Multi-agent Collaboration Framework for Long-horizon Planning

DGX agent

arXiv:2605.02168v1 Announce Type: cross Abstract: Language model (LM)-based agents have demonstrated promising capabilities in automating complex tasks from natural language instructions, yet they con

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Selector-Guided Autonomous Curriculum for One-Shot Reinforcement Learning from Verifiable Rewards

DGX agent

arXiv:2605.01823v1 Announce Type: new Abstract: Recently, Reinforcement Learning from Verifiable Rewards (RLVR) has been established as a highly effective technique for augmenting the math reasoning s

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

Towards High Fidelity Face Swapping: A Comprehensive Survey and New Benchmark

DGX agent

arXiv:2605.00883v1 Announce Type: new Abstract: Face swapping has witnessed significant progress in recent years, largely driven by advances in deep generative models such as GANs and diffusion models

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

VideoNet: A Large-Scale Dataset for Domain-Specific Action Recognition

DGX agent

arXiv:2605.02834v1 Announce Type: new Abstract: Videos are unique in their ability to capture actions which transcend multiple frames. Accordingly, for many years action recognition was the quintessen

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

VISTA: Video Interaction Spatio-Temporal Analysis Benchmark

DGX agent

arXiv:2605.01391v1 Announce Type: new Abstract: Existing benchmarks for Vision-Language Models (VLMs) primarily evaluate spatio-temporal understanding on simple single-action videos, closed attribute

model-releasesarxiv-cs-cv
5 May 2026
Model Releases

When Iterative RAG Beats Ideal Evidence: A Diagnostic Study in Scientific Multi-hop Question Answering

DGX agent

arXiv:2601.19827v3 Announce Type: replace Abstract: Retrieval-Augmented Generation (RAG) extends large language models (LLMs) beyond parametric knowledge, yet it is unclear when iterative retrieval-re

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

When RL Meets Adaptive Speculative Training: A Unified Training-Serving System

DGX agent

arXiv:2602.06932v3 Announce Type: replace Abstract: Speculative decoding can significantly accelerate LLM serving, yet most deployments today disentangle speculator training from serving, treating spe

model-releasesarxiv-cs-lg
5 May 2026
Model Releases

X2SAM: Any Segmentation in Images and Videos

DGX agent

arXiv:2605.00891v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have demonstrated strong image-level visual understanding and reasoning, yet their pixel-level perception acros

model-releasesarxiv-cs-cv
5 May 2026
Local Ai

Zero-Shot Confidence Estimation for Small LLMs: When Supervised Baselines Aren't Worth Training

DGX agent

arXiv:2605.02241v1 Announce Type: cross Abstract: How reliably can a small language model estimate its own correctness? The answer determines whether local-to-cloud routing-escalating queries a cheap

local-aiarxiv-cs-cl
5 May 2026
Model Releases

InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information

DGX agent

arXiv:2508.07630v2 Announce Type: replace Abstract: We introduce InterChart, a diagnostic benchmark that evaluates how well vision-language models (VLMs) reason across multiple related charts, a task

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering

DGX agent

arXiv:2505.23723v2 Announce Type: replace Abstract: The emergence of large language model (LLM)-based agents has significantly advanced the development of autonomous machine learning (ML) engineering.

model-releasesarxiv-cs-cl
4 May 2026
Research

The Algorithmic Gaze of Image Quality Assessment: An Audit and Trace Ethnography of the LAION-Aesthetics Predictor

DGX agent

arXiv:2601.09896v4 Announce Type: replace-cross Abstract: Visual generative AI models are trained using a one-size-fits-all measure of aesthetic appeal. However, what is deemed 'aesthetic' is inextric

researcharxiv-cs-cv
4 May 2026
Model Releases

ViLegalNLI: Natural Language Inference for Vietnamese Legal Texts

DGX agent

arXiv:2605.00116v1 Announce Type: new Abstract: In this article, we introduce ViLegalNLI, the first large-scale Vietnamese Natural Language Inference (NLI) dataset specifically constructed for the leg

model-releasesarxiv-cs-cl
4 May 2026
Tutorials

A generalised pre-training strategy for deep learning networks in semantic segmentation of remotely sensed images

DGX agent

arXiv:2604.27704v1 Announce Type: new Abstract: In the segmentation of remotely sensed images, deep learning models are typically pre-trained using large image databases like ImageNet before fine-tune

tutorialsarxiv-cs-cv
1 May 2026
Model Releases

AutoSP: Unlocking Long-Context LLM Training Via Compiler-Based Sequence Parallelism

DGX agent

arXiv:2604.27089v1 Announce Type: new Abstract: Large-language-models (LLMs) demonstrate enormous utility in long-context tasks which require processing prompts that consist of tens to hundreds of tho

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

CareGuardAI: Context-Aware Multi-Agent Guardrails for Clinical Safety & Hallucination Mitigation in Patient-Facing LLMs

DGX agent

arXiv:2604.26959v1 Announce Type: cross Abstract: Integrating large language models (LLMs) into patient-facing healthcare systems offers significant potential to improve access to medical information.

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Dynamic Scaled Gradient Descent for Stable Fine-Tuning for Classifications

DGX agent

arXiv:2604.27987v1 Announce Type: new Abstract: Fine-tuning pretrained models has become a standard approach to adapting pretrained knowledge to improve the accuracy on new sparse, imbalance datasets.

model-releasesarxiv-cs-lg
1 May 2026
Safety

Exploration Hacking: Can LLMs Learn to Resist RL Training?

DGX agent

arXiv:2604.28182v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become essential to the post-training of large language models (LLMs) for reasoning, agentic capabilities and alignmen

safetyarxiv-cs-cl
1 May 2026
Model Releases

Flattery in Motion: Benchmarking and Analyzing Sycophancy in Video-LLMs

DGX agent

arXiv:2506.07180v3 Announce Type: replace-cross Abstract: As video large language models (Video-LLMs) become increasingly integrated into real-world applications that demand grounded multimodal reason

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Mirage to Grounding: Towards Reliable Multimodal Circuit-to-Verilog Code Generation

DGX agent

arXiv:2604.27969v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) are increasingly used to translate visual artifacts into code, from UI mockups into HTML to scientific plots

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

From Unstructured Recall to Schema-Grounded Memory: Reliable AI Memory via Iterative, Schema-Aware Extraction

DGX agent

arXiv:2604.27906v1 Announce Type: new Abstract: Persistent AI memory is often reduced to a retrieval problem: store prior interactions as text, embed them, and ask the model to recover relevant contex

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Instruction Complexity Induces Positional Collapse in Adversarial LLM Evaluation

DGX agent

arXiv:2604.27249v1 Announce Type: cross Abstract: When instructed to underperform on multiple-choice evaluations, do language models engage with question content or fall back on positional shortcuts?

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

M-DaQ: Retrieving Samples with Multilingual Diversity and Quality for Instruction Fine-Tuning Datasets

DGX agent

arXiv:2509.15549v2 Announce Type: replace Abstract: Multilingual instruction fine-tuning (IFT) empowers large language models to generalize across diverse linguistic and cultural contexts; however, hi

model-releasesarxiv-cs-cl
1 May 2026
Model Releases

OR-VSKC: Resolving Visual-Semantic Knowledge Conflicts in Operating Rooms with Synthetic Data-Guided Alignment

DGX agent

arXiv:2506.22500v2 Announce Type: replace-cross Abstract: Automated identification of surgical safety risks is critical for improving patient outcomes; however, Multimodal Large Language Models (MLLMs

model-releasesarxiv-cs-ai
1 May 2026
Safety

Political Bias Audits of LLMs Capture Sycophancy to the Inferred Auditor

DGX agent

arXiv:2604.27633v1 Announce Type: new Abstract: Large language models (LLMs) are commonly evaluated for political bias based on their responses to fixed questionnaires, which typically place frontier

safetyarxiv-cs-ai
1 May 2026
Model Releases

PRISM: Pre-alignment via Black-box On-policy Distillation for Multimodal Reinforcement Learning

DGX agent

arXiv:2604.28123v1 Announce Type: cross Abstract: The standard post-training recipe for large multimodal models (LMMs) applies supervised fine-tuning (SFT) on curated demonstrations followed by reinfo

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Reduced NEXI protocol for the quantification of human gray matter microstructure on the Connectome 2.0 scanner

DGX agent

arXiv:2509.09513v2 Announce Type: replace-cross Abstract: Biophysical diffusion MRI models like Neurite Exchange Imaging (NEXI) are essential for probing gray matter microstructure, estimating compart

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

VeraRetouch: A Lightweight Fully Differentiable Framework for Multi-Task Reasoning Photo Retouching

DGX agent

arXiv:2604.27375v1 Announce Type: new Abstract: Reasoning photo retouching has gained significant traction, requiring models to analyze image defects, give reasoning processes, and execute precise ret

model-releasesarxiv-cs-cv
1 May 2026
Model Releases

When Roles Fail: Epistemic Constraints on Advocate Role Fidelity in LLM-Based Political Statement Analysis

DGX agent

arXiv:2604.27228v1 Announce Type: new Abstract: Democratic discourse analysis systems increasingly rely on multi-agent LLM pipelines in which distinct evaluator models are assigned adversarial roles t

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Allen AI just released the OlmPool research series on Hugging Face Early 7-8B checkpoints trained to 150B tokens exploring how minor archite…

DGX agent

Allen AI released the OlmPool research series on Hugging Face, featuring early 7-8B parameter language model checkpoints trained on 150 billion tokens. The research explores how minor architectural mo

model-releasesclem-delangue--x
30 Apr 2026
Model Releases

COP-GEN: Latent Diffusion Transformer for Copernicus Earth Observation Data

DGX agent

arXiv:2603.03239v2 Announce Type: replace Abstract: Earth observation applications increasingly rely on data from multiple sensors, including optical, radar, elevation, and land-cover. Relationships b

model-releasesarxiv-cs-cv
30 Apr 2026
Model Releases

DB-KSVD: Scalable Alternating Optimization for Disentangling High-Dimensional Embedding Spaces

DGX agent

arXiv:2505.18441v2 Announce Type: replace Abstract: Dictionary learning has recently emerged as a promising approach for mechanistic interpretability of large transformer models. Disentangling high-di

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

DSIPA: Detecting LLM-Generated Texts via Sentiment-Invariant Patterns Divergence Analysis

DGX agent

arXiv:2604.26328v1 Announce Type: cross Abstract: The rapid advancement of large language models (LLMs) presents new security challenges, particularly in detecting machine-generated text used for misi

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Efficient, VRAM-Constrained xLM Inference on Clients

DGX agent

arXiv:2604.26334v1 Announce Type: cross Abstract: To usher in the next round of client AI innovation, there is an urgent need to enable efficient, lossless inference of high-accuracy large language mo

model-releasesarxiv-cs-lg
30 Apr 2026
Model Releases

ELIQ: A Label-Free Framework for Quality Assessment of Evolving AI-Generated Images

DGX agent

arXiv:2602.03558v2 Announce Type: replace-cross Abstract: Generative text-to-image models are advancing at an unprecedented pace, continuously shifting the perceptual quality ceiling and rendering pre

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Enforcing Benign Trajectories: A Behavioral Firewall for Structured-Workflow AI Agents

DGX agent

arXiv:2604.26274v1 Announce Type: cross Abstract: Structured-workflow agents driven by large language models execute tool calls against sensitive external environments. We propose odename, a telemetry

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

EvoDev: An Iterative Feature-Driven Framework for End-to-End Software Development with LLM-based Agents

DGX agent

arXiv:2511.02399v2 Announce Type: replace-cross Abstract: Recent advances in large language model agents offer the promise of automating end-to-end software development from natural language requireme

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

HER: Human-like Reasoning and Reinforcement Learning for LLM Role-playing

DGX agent

arXiv:2601.21459v4 Announce Type: replace-cross Abstract: LLM role-playing, i.e., using LLMs to simulate specific personas, has emerged as a key capability in various applications, such as companionsh

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

MARVIS: Modality Adaptive Reasoning over VISualizations

DGX agent

arXiv:2507.01544v2 Announce Type: replace Abstract: Predictive applications of machine learning often rely on small (sub 1 Bn parameter) specialized models tuned to particular domains or modalities. S

model-releasesarxiv-cs-lg
30 Apr 2026
← Previous
1…419420421422423…1371
Next →