AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,766 results
Model Releases

'The Whole Is Greater Than the Sum of Its Parts': A Compatibility-Aware Multi-Teacher CoT Distillation Framework

DGX agent

arXiv:2601.13992v2 Announce Type: replace-cross Abstract: Chain-of-Thought (CoT) reasoning empowers Large Language Models (LLMs) with remarkable capabilities but typically requires prohibitive paramet

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

TPV: Parameter Perturbations Through the Lens of Test Prediction Variance

DGX agent

arXiv:2512.11089v4 Announce Type: replace-cross Abstract: We introduce test prediction variance (TPV)--the first-order sensitivity of a trained model's outputs to parameter perturbations--as a unifyin

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

UCSF-PDGM-VQA: Visual Question Answering dataset for brain tumor MRI interpretation

DGX agent

arXiv:2605.17140v1 Announce Type: cross Abstract: Brain tumor diagnosis is largely dependent on Magnetic Resonance Imaging (MRI) evaluation, which requires radiologists to synthesize thousands of imag

model-releasesarxiv-cs-ai
19 May 2026
Safety

Weak-to-Strong Elicitation via Mismatched Wrong Drafts

DGX agent

arXiv:2605.17314v1 Announce Type: cross Abstract: We consider whether off-policy experience from a smaller, weaker model can elicit capability in a stronger learner that on-policy RL fine-tuning (e.g.

safetyarxiv-cs-ai
19 May 2026
Research

When Efficiency Backfires: Cascading LLMs Trigger Cascade Failure under Adversarial Attack

DGX agent

arXiv:2605.17288v1 Announce Type: cross Abstract: Large Language Model (LLM) cascade systems are designed to balance efficiency and performance by processing queries with lightweight models while sele

researcharxiv-cs-ai
19 May 2026
Applications

XCTFormer: Leveraging Cross-Channel and Cross-Time Dependencies for Enhanced Time-Series Analysis

DGX agent

arXiv:2605.18534v1 Announce Type: new Abstract: Multivariate time-series analysis involves extracting informative representations from sequences of multiple interdependent variables, supporting tasks

applicationsarxiv-cs-lg
19 May 2026
Safety

ZeroSiam: An Efficient Asymmetry for Test-Time Entropy Optimization without Collapse

DGX agent

arXiv:2509.23183v3 Announce Type: replace Abstract: Test-time entropy minimization helps adapt a model to novel environments and incentivize its reasoning capability, unleashing the model's potential

safetyarxiv-cs-lg
19 May 2026
Model Releases

A Reproducible and Physically Feasible Dynamic Parameter Identification Framework for a Low-Cost Robot Arm

DGX agent

arXiv:2605.15949v1 Announce Type: new Abstract: This paper presents a reproducible and physically feasible dynamic parameter identification framework for CRANE-X7, a low-cost robot arm driven by modul

model-releasesarxiv-cs-ro
18 May 2026
Research

A Unified Perturbation Framework for Analyzing Leaderboard Stability and Manipulation

DGX agent

arXiv:2605.15761v1 Announce Type: new Abstract: Evaluation leaderboards such as LMArena play a central role in benchmarking large language models by aggregating pairwise human preferences into model r

researcharxiv-cs-lg
18 May 2026
Model Releases

Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design

DGX agent

arXiv:2605.15871v1 Announce Type: new Abstract: Toward recursive self-improvement, we investigate LLM agents autonomously designing foundation models beyond standard Transformers. We introduce a dual-

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Deep Pre-Alignment for VLMs

DGX agent

arXiv:2605.15300v1 Announce Type: new Abstract: Most Vision Language Models (VLMs) directly map outputs from ViT encoders to the LLM via a lightweight projector. While effective, recent analysis sugge

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

DimMem: Dimensional Structuring for Efficient Long-Term Agent Memory

DGX agent

arXiv:2605.15759v1 Announce Type: new Abstract: Large language model (LLM) agents require long-term memory to leverage information from past interactions. However, existing memory systems often face a

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Federated Learning of Spiking Neural Networks under Heterogeneous Temporal Resolutions

DGX agent

arXiv:2605.15355v1 Announce Type: new Abstract: Spiking neural networks (SNNs) are biologically inspired energy-efficient models that use sparse binary spike-based communication between neurons, makin

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Flash-GRPO: Efficient Alignment for Video Diffusion via One-Step Policy Optimization

DGX agent

arXiv:2605.15980v1 Announce Type: new Abstract: Group Relative Policy Optimization has emerged as essential for aligning video diffusion models with human preferences, but faces a critical computation

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

FORGE: Self-Evolving Agent Memory With No Weight Updates via Population Broadcast

DGX agent

arXiv:2605.16233v1 Announce Type: new Abstract: Can LLM agents improve decision-making through self-generated memory without gradient updates? We propose FORGE (Failure-Optimized Reflective Graduation

model-releasesarxiv-cs-ai
18 May 2026
Applications

Fortress: A Case Study in Stabilizing Search Recommendations via Temporal Data Augmentation and Feature Pruning

DGX agent

arXiv:2605.15299v1 Announce Type: cross Abstract: In search and recommendation systems, predictive models often suffer from temporal instability when certain input features introduce volatility in out

applicationsarxiv-cs-ai
18 May 2026
Local Ai

Grounded Reinforcement Learning for Visual Reasoning

DGX agent

arXiv:2505.23678v3 Announce Type: replace Abstract: While reinforcement learning (RL) over chains of thought has significantly advanced language models in tasks such as mathematics and coding, visual

local-aiarxiv-cs-cv
18 May 2026
Tutorials

How to Choose Your Teacher for Fine Grained Image Recognition

DGX agent

arXiv:2605.15689v1 Announce Type: new Abstract: Fine-grained image recognition classifies subcategories such as bird species or car models. While state-of-the-art (SOTA) models are accurate, they are

tutorialsarxiv-cs-cv
18 May 2026
Research

LPDS: Evaluating LLM Robustness Through Logic-Preserving Difficulty Scaling

DGX agent

arXiv:2605.15393v1 Announce Type: new Abstract: As large language models (LLMs) are increasingly deployed to perform tasks with minimal human oversight, it is crucial that these models operate robustl

researcharxiv-cs-lg
18 May 2026
Model Releases

Optimizing LLM Inference: Fluid-Guided Online Scheduling with Memory Constraints

DGX agent

arXiv:2504.11320v3 Announce Type: replace-cross Abstract: Large language models now serve millions of users daily, with providers incurring costs exceeding $700,000 per day. Each request requires toke

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

PerfCodeBench: Benchmarking LLMs for System-Level High-Performance Code Optimization

DGX agent

arXiv:2605.15222v1 Announce Type: cross Abstract: Large language models (LLMs) can often generate functionally correct code, but their ability to produce efficient implementations for performance-crit

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Representation Without Reward: A JEPA Audit for LLM Fine-Tuning

DGX agent

arXiv:2605.15394v1 Announce Type: cross Abstract: Joint-embedding predictive architectures (JEPAs) propose that a model should learn more useful abstractions when trained to predict latent representat

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces

DGX agent

arXiv:2605.15215v1 Announce Type: new Abstract: Recently, skills have been widely adopted in large language model (LLM)-based agent systems across various domains. In existing frameworks, skills are t

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

I like what Langchain recently released in their deepagents harness, which is an adapter to modify the syntax of primitive file system comma…

DGX agent

I like what Langchain recently released in their deepagents harness, which is an adapter to modify the syntax of primitive file system commands depending on the model Claude likes “Bash”, Gemini likes

model-releasesharrison-chase--x
16 May 2026
Model Releases

Cattle Trade: A Multi-Agent Benchmark for LLM Bluffing, Bidding, and Bargaining

DGX agent

arXiv:2605.14537v1 Announce Type: new Abstract: We introduce extsc{Cattle Trade, a multi-agent benchmark for evaluating large language models (LLMs) as agents in strategic reasoning under imperfect in

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

ClawForge: Generating Executable Interactive Benchmarks for Command-Line Agents

DGX agent

arXiv:2605.14133v1 Announce Type: new Abstract: Interactive agent benchmarks face a tension between scalable construction and realistic workflow evaluation. Hand-authored tasks are expensive to extend

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Communication-Efficient Federated Fine-Tuning

DGX agent

arXiv:2505.04535v3 Announce Type: replace Abstract: Federated Learning (FL) enables the utilization of vast, previously inaccessible data sources. At the same time, pre-trained Language Models (LMs) h

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)

DGX agent

arXiv:2511.13397v2 Announce Type: replace-cross Abstract: The remarkable progress of Vision-Language Models (VLMs) on a variety of tasks has raised interest in their application to automated driving.

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Do-Undo Bench: Reversibility for Action Understanding in Image Generation

DGX agent

arXiv:2512.13609v2 Announce Type: replace Abstract: We introduce the Do-Undo task and benchmark to address a critical gap in vision-language models: understanding and generating plausible scene transf

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

EndPrompt: Efficient Long-Context Extension via Terminal Anchoring

DGX agent

arXiv:2605.14589v1 Announce Type: new Abstract: Extending the context window of large language models typically requires training on sequences at the target length, incurring quadratic memory and comp

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents

DGX agent

arXiv:2605.14153v1 Announce Type: cross Abstract: Exploitation is not a binary event. It is a ladder of acquiring progressive capabilities, from executing a single buggy line of code to taking full co

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Forgetting That Sticks: Quantization-Permanent Unlearning via Circuit Attribution

DGX agent

arXiv:2605.15138v1 Announce Type: cross Abstract: Standard unlearning evaluations measure behavioral suppression in full precision, immediately after training, despite every deployed language model be

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

HDRFace: Rethinking Face Restoration with High-Dimensional Representation

DGX agent

arXiv:2605.14821v1 Announce Type: new Abstract: Face restoration under complex degradations still remains an ill-posed inverse problem due to severe information loss. Although diffusion models benefit

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

HiSem: Hierarchical Semantic Disentangling for Remote Sensing Image Change Captioning

DGX agent

arXiv:2605.15024v1 Announce Type: new Abstract: Remote sensing image change captioning (RSICC) aims to achieve high-level semantic understanding of genuine changes occurring between bi-temporal images

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

LLMs Know When They Know, but Do Not Act on It: A Metacognitive Harness for Test-time Scaling

DGX agent

arXiv:2605.14186v1 Announce Type: new Abstract: Large language models (LLMs) often expose useful signals of self-monitoring: before solving a problem, they can estimate whether they are likely to succ

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

NodeSynth: Socially Aligned Synthetic Data for AI Evaluation

DGX agent

arXiv:2605.14381v1 Announce Type: cross Abstract: Recent advancements in generative AI facilitate large-scale synthetic data generation for model evaluation. However, without targeted approaches, thes

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

Test-Time Learning with an Evolving Library

DGX agent

arXiv:2605.14477v1 Announce Type: new Abstract: We introduce EvoLib, a test-time learning framework that enables large language models to accumulate, reuse, and evolve knowledge across problem instanc

model-releasesarxiv-cs-lg
15 May 2026
Model Releases

Text Knows What, Tables Know When: Clinical Timeline Reconstruction via Retrieval-Augmented Multimodal Alignment

DGX agent

arXiv:2605.15168v1 Announce Type: cross Abstract: Reconstructing precise clinical timelines is essential for modeling patient trajectories and forecasting risk in complex, heterogeneous conditions lik

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

TFGN: Task-Free, Replay-Free Continual Pre-Training Without Catastrophic Forgetting at LLM Scale

DGX agent

arXiv:2605.15053v1 Announce Type: cross Abstract: Continually pre-training a large language model on heterogeneous text domains, without replay or task labels, has remained an unsolved architectural p

model-releasesarxiv-cs-ai
15 May 2026
Applications

What Makes Words Hard? Sakura at BEA 2026 Shared Task on Vocabulary Difficulty Prediction

DGX agent

arXiv:2605.14257v1 Announce Type: new Abstract: We describe two types of models for vocabulary difficulty prediction: a high-accuracy black-box model, which achieved the top shared task result in the

applicationsarxiv-cs-cl
15 May 2026
Tutorials

Characteristic Root Analysis and Regularization for Linear Time Series Forecasting

DGX agent

arXiv:2509.23597v5 Announce Type: replace-cross Abstract: Time series forecasting remains a critical challenge across numerous domains, yet the effectiveness of complex models often varies unpredictab

tutorialsarxiv-cs-ai
14 May 2026
Model Releases

Controlling Logical Collapse in LLMs via Algebraic Ontology Projection over F2

DGX agent

arXiv:2605.12968v1 Announce Type: cross Abstract: Do large language models internally encode ontological relations in a formally verifiable algebraic structure? We introduce Algebraic Ontology Project

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

CR-Net: Scaling Parameter-Efficient Training with Cross-Layer Low-Rank Structure

DGX agent

arXiv:2509.18993v3 Announce Type: replace Abstract: Low-rank architectures have become increasingly important for efficient large language model (LLM) pre-training, providing substantial reductions in

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

DocAtlas: Multilingual Document Understanding Across 80+ Languages

DGX agent

arXiv:2605.12623v1 Announce Type: cross Abstract: Multilingual document understanding remains limited for low-resource languages due to scarce training data and model-based annotation pipelines that p

model-releasesarxiv-cs-cv
14 May 2026
Model Releases

Efficient compression of neural networks and datasets

DGX agent

arXiv:2505.17469v2 Announce Type: replace-cross Abstract: Compression and generalization are fundamentally related through Solomonoff induction and the minimum description length principle (MDL), whic

model-releasesarxiv-cs-ai
14 May 2026
Tutorials

How Do Transformers Learn to Associate Tokens: Gradient Leading Terms Bring Mechanistic Interpretability

DGX agent

arXiv:2601.19208v2 Announce Type: replace-cross Abstract: Semantic associations such as the link between 'bird' and 'flew' are foundational for language modeling as they enable models to go beyond mem

tutorialsarxiv-cs-lg
14 May 2026
Model Releases

Identifying the nonlinear string dynamics with port-Hamiltonian neural networks

DGX agent

arXiv:2605.12785v1 Announce Type: new Abstract: Hybrid machine learning combines physical knowledge with data-driven models to enhance interpretability and performance. In this context, Port-Hamiltoni

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

LoRA-Mixer: Coordinate Modular LoRA Experts Through Serial Attention Routing

DGX agent

arXiv:2507.00029v2 Announce Type: replace-cross Abstract: Recent attempts to combine low-rank adaptation (LoRA) with mixture-of-experts (MoE) for multi-task adaptation of Large Language Models (LLMs)

model-releasesarxiv-cs-ai
14 May 2026
← Previous
1…415416417418419…1371
Next →